Skip to content

ML Engineer, Foundation Model Evaluation

Is this job for you?

Build your CV and see how well you match this role — and every other one.

Build my CV

The role

Develop and extend cutting-edge ML and robotics research to advance evaluation methodologies for embodied AI agents and foundation models.
Define benchmarks, metrics and evaluation protocols to assess quality, safety and realism.
Build and maintain large-scale data and evaluation pipelines to enable robust ML model assessment.
Collaborate across teams to land disruptive evaluation tech into production and work with state-of-the-art foundation models.
Proficiency in Python and modern deep learning frameworks (PyTorch, JAX, TensorFlow); strong software engineering; C++ preferred.
Hybrid remote work in the United States with a track record in ML model evaluation or related research.

See the full job post

Responsibilities, requirements, skills and benefits — create your free account.

At least 8 characters, one uppercase letter and one digit.

Already have an account? Log in

Similar openings

Other roles that could suit you.

See all →

Your location

Jobs and companies will be filtered on this country.

Suggested

All countries 65