Software Engineer/Data Scientist, Large Model Evaluation
职位介绍
Join the Large Model Evaluation team to advance driving ML model evaluation through novel metrics, sampling methods, and simulation-based assessments.
Develop metrics and sampling strategies to quantify trajectories generated by ML models.
Design data pipelines for signal discovery, labeling, feature extraction, and metric computation at scale.
Analyze data to diagnose regressions and translate findings into actionable model improvements.
Create creative simulation strategies to evaluate the driving performance of generative AI models and uncover edge cases.
Collaborate with software engineers and researchers to ensure rigorous, safe, and scalable model evaluation.
Develop metrics and sampling strategies to quantify trajectories generated by ML models.
Design data pipelines for signal discovery, labeling, feature extraction, and metric computation at scale.
Analyze data to diagnose regressions and translate findings into actionable model improvements.
Create creative simulation strategies to evaluate the driving performance of generative AI models and uncover edge cases.
Collaborate with software engineers and researchers to ensure rigorous, safe, and scalable model evaluation.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
已有账户? 登录
相似职位
其他可能适合您的职位。
远程办公partial
城市San Francisco, 美国