Skip to content

Senior Software Engineer/Data Scientist, Large Model Evaluation

Is this job for you?

Build your CV and see how well you match this role — and every other one.

Build my CV

The role

Join the Large Model Evaluation team to design and implement novel evaluation metrics and sampling strategies for driving trajectories produced by ML models. Use creative simulation approaches to quantify driving performance of generative AI systems, identify edge cases, and deliver reliable performance insights to guide model development and deployment. Build end-to-end data pipelines for signal discovery, labeling, feature extraction, and metric computation from large-scale simulations; perform data analysis to diagnose regressions and root causes. Collaborate with software engineers and researchers to scale evaluation tools across large ML models and real-world driving contexts. Requirements include 5+ years in quantitative software engineering, proficiency in Python or C++, experience building data pipelines, system evaluation, and productionized tooling; knowledge of transformer architectures and ML fundamentals; familiarity with frameworks such as JAX or TensorFlow is a plus; experience with simulation, robotics, or autonomous vehicles is preferred.

See the full job post

Responsibilities, requirements, skills and benefits — create your free account.

At least 8 characters, one uppercase letter and one digit.

Already have an account? Log in

Similar openings

Other roles that could suit you.

See all →

Your location

Jobs and companies will be filtered on this country.

Suggested

All countries 65