ML Engineer, Foundation Model Evaluation
¿Es esta oferta para usted?
Crear mi CV Cree su CV y descubra su porcentaje de coincidencia con este puesto — y con todos los demás.
El puesto
Develop and extend cutting-edge ML and robotics research to advance evaluation methodologies for embodied AI agents and foundation models.
Define benchmarks, metrics and evaluation protocols to assess quality, safety and realism.
Build and maintain large-scale data and evaluation pipelines to enable robust ML model assessment.
Collaborate across teams to land disruptive evaluation tech into production and work with state-of-the-art foundation models.
Proficiency in Python and modern deep learning frameworks (PyTorch, JAX, TensorFlow); strong software engineering; C++ preferred.
Hybrid remote work in the United States with a track record in ML model evaluation or related research.
Define benchmarks, metrics and evaluation protocols to assess quality, safety and realism.
Build and maintain large-scale data and evaluation pipelines to enable robust ML model assessment.
Collaborate across teams to land disruptive evaluation tech into production and work with state-of-the-art foundation models.
Proficiency in Python and modern deep learning frameworks (PyTorch, JAX, TensorFlow); strong software engineering; C++ preferred.
Hybrid remote work in the United States with a track record in ML model evaluation or related research.
Ver la oferta completa
Funciones, perfil, competencias y ventajas — crea tu cuenta gratis.
¿Ya tiene una cuenta? Iniciar sesión
Ofertas similares
Otros puestos que podrían encajar.
TeletrabajoPartial remote
CiudadRemote / Hybrid (US), United States