Software Engineer/Data Scientist, Large Model Evaluation
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Join the Large Model Evaluation team to advance driving ML model evaluation through novel metrics, sampling methods, and simulation-based assessments.
Develop metrics and sampling strategies to quantify trajectories generated by ML models.
Design data pipelines for signal discovery, labeling, feature extraction, and metric computation at scale.
Analyze data to diagnose regressions and translate findings into actionable model improvements.
Create creative simulation strategies to evaluate the driving performance of generative AI models and uncover edge cases.
Collaborate with software engineers and researchers to ensure rigorous, safe, and scalable model evaluation.
Develop metrics and sampling strategies to quantify trajectories generated by ML models.
Design data pipelines for signal discovery, labeling, feature extraction, and metric computation at scale.
Analyze data to diagnose regressions and translate findings into actionable model improvements.
Create creative simulation strategies to evaluate the driving performance of generative AI models and uncover edge cases.
Collaborate with software engineers and researchers to ensure rigorous, safe, and scalable model evaluation.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
Already have an account? Log in
Similar openings
Other roles that could suit you.
Remote workpartial
CitySan Francisco, United States