Software Engineer/Data Scientist, Large Model Evaluation
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Develop novel metrics and sampling techniques to measure driving trajectories generated by ML models.
Employ creative simulation strategies to evaluate the driving performance of generative AI models and identify edge cases.
Build scalable data pipelines for signal discovery, labeling, feature extraction, and metric computation from large-scale simulations.
Conduct data analysis to diagnose regressions and inform model improvements.
Collaborate with world-class engineering and research teams to advance evaluation methodologies.
Drive rigorous, safety-focused assessment of deployed models in real-world driving contexts.
Employ creative simulation strategies to evaluate the driving performance of generative AI models and identify edge cases.
Build scalable data pipelines for signal discovery, labeling, feature extraction, and metric computation from large-scale simulations.
Conduct data analysis to diagnose regressions and inform model improvements.
Collaborate with world-class engineering and research teams to advance evaluation methodologies.
Drive rigorous, safety-focused assessment of deployed models in real-world driving contexts.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inSimilar openings
Other roles that could suit you.
Remote workno
CitySan Francisco, United States