Skip to content

Staff Software Engineer/Data Scientist, Large Model Evaluation

Is this job for you?

Build your CV and see how well you match this role — and every other one.

Build my CV

The role

Join a dedicated evaluation team to define novel metrics and sampling methods for driving trajectories produced by ML models. Develop simulation-based evaluation strategies to measure the driving performance of generative AI systems. Build scalable data pipelines for signal discovery, data labeling, feature extraction, and metric computation from large-scale simulations. Conduct data analysis to diagnose regressions and provide reliable performance insights that inform model development and deployment. Collaborate with engineering and research teams to translate quantitative findings into production-ready tools. Required: 7+ years of quantitative software engineering, strong Python or C++, knowledge of AI fundamentals and end-to-end model evaluation.

See the full job post

Responsibilities, requirements, skills and benefits — create your free account.

At least 8 characters, one uppercase letter and one digit.

Already have an account? Log in

Similar openings

Other roles that could suit you.

See all →

Your location

Jobs and companies will be filtered on this country.

Suggested

All countries 65