跳到正文

Machine Learning Engineer, Reinforcement Learning & Reward Modeling

这个职位适合你吗?

创建简历,即可看到你与这个职位——以及其他所有职位——的匹配度。

创建我的简历

职位介绍

Design and optimize end-to-end pipelines for training reward models and RL agents to be reproducible and high-throughput.
Develop tooling for data processing, annotation, and inference within RL workflows.
Build, refine, and deploy reward models that encode safe, interpretable, and effective driving behaviours.
Integrate reward models with diverse data sources: real-world trajectories, simulation, and synthetic datasets.
Conduct ablations, hyperparameter explorations, and controlled studies to analyse how reward structures and training dynamics affect policy performance.
Diagnose failure modes, iterate on reward objectives, and partner with RL scientists to translate ideas into scalable engineering solutions and testing frameworks.

查看完整职位

工作职责、任职要求、技能与福利 — 免费创建账号即可查看。

至少 8 个字符,需包含一个大写字母和一个数字。

已有账户? 登录

相似职位

其他可能适合您的职位。

查看全部 →

您的城市

职位和企业将按该国家筛选。

推荐

所有国家 66