Staff Software Engineer/Data Scientist, Large Model Evaluation
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Develop novel metrics and sampling techniques to measure driving trajectories generated by ML models.
Design and execute creative simulation strategies to assess driving performance of generative AI models and identify edge cases.
Build data pipelines for signal discovery, data labeling, feature extraction, and metric computation from large-scale simulations.
Perform data analysis to diagnose regressions and translate quantitative findings into production-ready tools.
Collaborate with software and research teams to align evaluation methods with model development and deployment goals.
Operate in a fast-paced, safety-critical environment where rigorous measurement guides deployment decisions.
Design and execute creative simulation strategies to assess driving performance of generative AI models and identify edge cases.
Build data pipelines for signal discovery, data labeling, feature extraction, and metric computation from large-scale simulations.
Perform data analysis to diagnose regressions and translate quantitative findings into production-ready tools.
Collaborate with software and research teams to align evaluation methods with model development and deployment goals.
Operate in a fast-paced, safety-critical environment where rigorous measurement guides deployment decisions.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inSimilar openings
Other roles that could suit you.
Remote workPartial
CitySan Francisco, United States