跳到正文

Senior Machine Learning Engineer - Evaluation

这个职位适合你吗?

创建简历,即可看到你与这个职位——以及其他所有职位——的匹配度。

创建我的简历

职位介绍

Lead the design and implementation of offline evaluation pipelines for embodied AI models, building scalable and interpretable benchmarks that cover action, perception, and language-grounded reasoning. Design metrics across vision, language, and driving tasks to robustly evaluate model behavior and inform deployment readiness. Drive human annotation workflows, including task design, QA, and coordination with internal teams and external partners, to ensure high-quality ground-truth data. Collaborate with science, datasets, and infrastructure teams to align evaluation with product goals and system safety. Analyze offline metrics and correlate them with online performance to guide model selection and risk assessment. Demonstrate strong Python software engineering and data processing skills to deliver scalable, reproducible evaluation tooling.

查看完整职位

工作职责、任职要求、技能与福利 — 免费创建账号即可查看。

至少 6 个字符。越长越安全。

已有账户?

相似职位

其他可能适合您的职位。

查看全部 →

您的城市

职位和企业将按该国家筛选。

推荐

所有国家 66