Skip to content

Software Engineer/Data Scientist, Large Model Evaluation

Is this job for you?

Build your CV and see how well you match this role — and every other one.

Build my CV

The role

Join the Large Model Evaluation team to advance driving ML model evaluation through novel metrics, sampling methods, and simulation-based assessments.
Develop metrics and sampling strategies to quantify trajectories generated by ML models.
Design data pipelines for signal discovery, labeling, feature extraction, and metric computation at scale.
Analyze data to diagnose regressions and translate findings into actionable model improvements.
Create creative simulation strategies to evaluate the driving performance of generative AI models and uncover edge cases.
Collaborate with software engineers and researchers to ensure rigorous, safe, and scalable model evaluation.

See the full job post

Responsibilities, requirements, skills and benefits — create your free account.

At least 8 characters, one uppercase letter and one digit.

Already have an account? Log in

Similar openings

Other roles that could suit you.

See all →

Your location

Jobs and companies will be filtered on this country.

Suggested

All countries 65