ML Engineer, Foundation Model Evaluation
仕事内容
Develop and extend cutting-edge ML and robotics research to advance evaluation methodologies for embodied AI agents and foundation models.
Define benchmarks, metrics and evaluation protocols to assess quality, safety and realism.
Build and maintain large-scale data and evaluation pipelines to enable robust ML model assessment.
Collaborate across teams to land disruptive evaluation tech into production and work with state-of-the-art foundation models.
Proficiency in Python and modern deep learning frameworks (PyTorch, JAX, TensorFlow); strong software engineering; C++ preferred.
Hybrid remote work in the United States with a track record in ML model evaluation or related research.
Define benchmarks, metrics and evaluation protocols to assess quality, safety and realism.
Build and maintain large-scale data and evaluation pipelines to enable robust ML model assessment.
Collaborate across teams to land disruptive evaluation tech into production and work with state-of-the-art foundation models.
Proficiency in Python and modern deep learning frameworks (PyTorch, JAX, TensorFlow); strong software engineering; C++ preferred.
Hybrid remote work in the United States with a track record in ML model evaluation or related research.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログイン類似の求人
あなたに合いそうな他の職種。
リモートワークPartial remote
勤務地Remote / Hybrid (US), United States