Machine Learning Engineer, Reinforcement Learning & Reward Modeling
¿Es esta oferta para usted?
Crear mi CV Cree su CV y descubra su porcentaje de coincidencia con este puesto — y con todos los demás.
Sigue tus candidaturas desde el móvil La app gratuita de Whileresume, en iPhone y Android.
El puesto
Design and optimize end-to-end pipelines for training reward models and RL agents to be reproducible and high-throughput.
Develop tooling for data processing, annotation, and inference within RL workflows.
Build, refine, and deploy reward models that encode safe, interpretable, and effective driving behaviours.
Integrate reward models with diverse data sources: real-world trajectories, simulation, and synthetic datasets.
Conduct ablations, hyperparameter explorations, and controlled studies to analyse how reward structures and training dynamics affect policy performance.
Diagnose failure modes, iterate on reward objectives, and partner with RL scientists to translate ideas into scalable engineering solutions and testing frameworks.
Develop tooling for data processing, annotation, and inference within RL workflows.
Build, refine, and deploy reward models that encode safe, interpretable, and effective driving behaviours.
Integrate reward models with diverse data sources: real-world trajectories, simulation, and synthetic datasets.
Conduct ablations, hyperparameter explorations, and controlled studies to analyse how reward structures and training dynamics affect policy performance.
Diagnose failure modes, iterate on reward objectives, and partner with RL scientists to translate ideas into scalable engineering solutions and testing frameworks.
Ver la oferta completa
Funciones, perfil, competencias y ventajas — crea tu cuenta gratis.
o
¿Ya tiene una cuenta?
Iniciar sesiónEstas ofertas podrían interesarte
Ninguna oferta realmente parecida por ahora — estas son las más recientes.