Staff Software Engineer/Data Scientist, Large Model Evaluation
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Join a dedicated evaluation team to define novel metrics and sampling methods for driving trajectories produced by ML models. Develop simulation-based evaluation strategies to measure the driving performance of generative AI systems. Build scalable data pipelines for signal discovery, data labeling, feature extraction, and metric computation from large-scale simulations. Conduct data analysis to diagnose regressions and provide reliable performance insights that inform model development and deployment. Collaborate with engineering and research teams to translate quantitative findings into production-ready tools. Required: 7+ years of quantitative software engineering, strong Python or C++, knowledge of AI fundamentals and end-to-end model evaluation.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
Already have an account? Log in
Similar openings
Other roles that could suit you.
Remote workPartial
CitySan Francisco, United States