ML Engineer, Foundation Model Evaluation
职位介绍
Develop and extend cutting-edge ML and robotics research to advance evaluation methodologies for embodied AI agents and foundation models.
Define benchmarks, metrics and evaluation protocols to assess quality, safety and realism.
Build and maintain large-scale data and evaluation pipelines to enable robust ML model assessment.
Collaborate across teams to land disruptive evaluation tech into production and work with state-of-the-art foundation models.
Proficiency in Python and modern deep learning frameworks (PyTorch, JAX, TensorFlow); strong software engineering; C++ preferred.
Hybrid remote work in the United States with a track record in ML model evaluation or related research.
Define benchmarks, metrics and evaluation protocols to assess quality, safety and realism.
Build and maintain large-scale data and evaluation pipelines to enable robust ML model assessment.
Collaborate across teams to land disruptive evaluation tech into production and work with state-of-the-art foundation models.
Proficiency in Python and modern deep learning frameworks (PyTorch, JAX, TensorFlow); strong software engineering; C++ preferred.
Hybrid remote work in the United States with a track record in ML model evaluation or related research.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录相似职位
其他可能适合您的职位。
远程办公Partial remote
城市Remote / Hybrid (US), United States