Machine Learning Engineer, Reinforcement Learning & Reward Modeling
応募状況をスマホで確認 Whileresume の無料アプリ(iPhone・Android)。
仕事内容
Design and optimize end-to-end pipelines for training reward models and RL agents to be reproducible and high-throughput.
Develop tooling for data processing, annotation, and inference within RL workflows.
Build, refine, and deploy reward models that encode safe, interpretable, and effective driving behaviours.
Integrate reward models with diverse data sources: real-world trajectories, simulation, and synthetic datasets.
Conduct ablations, hyperparameter explorations, and controlled studies to analyse how reward structures and training dynamics affect policy performance.
Diagnose failure modes, iterate on reward objectives, and partner with RL scientists to translate ideas into scalable engineering solutions and testing frameworks.
Develop tooling for data processing, annotation, and inference within RL workflows.
Build, refine, and deploy reward models that encode safe, interpretable, and effective driving behaviours.
Integrate reward models with diverse data sources: real-world trajectories, simulation, and synthetic datasets.
Conduct ablations, hyperparameter explorations, and controlled studies to analyse how reward structures and training dynamics affect policy performance.
Diagnose failure modes, iterate on reward objectives, and partner with RL scientists to translate ideas into scalable engineering solutions and testing frameworks.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログインこちらの求人もおすすめです
近い求人はまだありません — 最新の求人をご紹介します。