ML Engineer, Foundation Model Evaluation
Cette offre est-elle pour vous ?
Créer mon CV Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.
Le poste
Develop and extend cutting-edge ML and robotics research to advance evaluation methodologies for embodied AI agents and foundation models.
Define benchmarks, metrics and evaluation protocols to assess quality, safety and realism.
Build and maintain large-scale data and evaluation pipelines to enable robust ML model assessment.
Collaborate across teams to land disruptive evaluation tech into production and work with state-of-the-art foundation models.
Proficiency in Python and modern deep learning frameworks (PyTorch, JAX, TensorFlow); strong software engineering; C++ preferred.
Hybrid remote work in the United States with a track record in ML model evaluation or related research.
Define benchmarks, metrics and evaluation protocols to assess quality, safety and realism.
Build and maintain large-scale data and evaluation pipelines to enable robust ML model assessment.
Collaborate across teams to land disruptive evaluation tech into production and work with state-of-the-art foundation models.
Proficiency in Python and modern deep learning frameworks (PyTorch, JAX, TensorFlow); strong software engineering; C++ preferred.
Hybrid remote work in the United States with a track record in ML model evaluation or related research.
Voir l'offre en entier
Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.
Déjà un compte ? Se connecter
Offres similaires
D'autres postes qui pourraient vous convenir.
TélétravailPartial remote
VilleRemote / Hybrid (US), United States