Software Engineer/Data Scientist, Large Model Evaluation
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
Track your applications on mobile The free Whileresume app, on iPhone and Android.
The role
Join the Large Model Evaluation team to advance driving ML model evaluation through novel metrics, sampling methods, and simulation-based assessments.
Develop metrics and sampling strategies to quantify trajectories generated by ML models.
Design data pipelines for signal discovery, labeling, feature extraction, and metric computation at scale.
Analyze data to diagnose regressions and translate findings into actionable model improvements.
Create creative simulation strategies to evaluate the driving performance of generative AI models and uncover edge cases.
Collaborate with software engineers and researchers to ensure rigorous, safe, and scalable model evaluation.
Develop metrics and sampling strategies to quantify trajectories generated by ML models.
Design data pipelines for signal discovery, labeling, feature extraction, and metric computation at scale.
Analyze data to diagnose regressions and translate findings into actionable model improvements.
Create creative simulation strategies to evaluate the driving performance of generative AI models and uncover edge cases.
Collaborate with software engineers and researchers to ensure rigorous, safe, and scalable model evaluation.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inYou might also like these jobs
No closely matching jobs yet — here are the most recent ones.