Machine Learning Engineer, Reinforcement Learning & Reward Modeling
Ist diese Stelle etwas für Sie?
Lebenslauf erstellen Erstellen Sie Ihren Lebenslauf und entdecken Sie Ihre Übereinstimmung mit dieser Stelle — und mit allen anderen.
Die Stelle
Design and optimize end-to-end pipelines for training reward models and RL agents to be reproducible and high-throughput.
Develop tooling for data processing, annotation, and inference within RL workflows.
Build, refine, and deploy reward models that encode safe, interpretable, and effective driving behaviours.
Integrate reward models with diverse data sources: real-world trajectories, simulation, and synthetic datasets.
Conduct ablations, hyperparameter explorations, and controlled studies to analyse how reward structures and training dynamics affect policy performance.
Diagnose failure modes, iterate on reward objectives, and partner with RL scientists to translate ideas into scalable engineering solutions and testing frameworks.
Develop tooling for data processing, annotation, and inference within RL workflows.
Build, refine, and deploy reward models that encode safe, interpretable, and effective driving behaviours.
Integrate reward models with diverse data sources: real-world trajectories, simulation, and synthetic datasets.
Conduct ablations, hyperparameter explorations, and controlled studies to analyse how reward structures and training dynamics affect policy performance.
Diagnose failure modes, iterate on reward objectives, and partner with RL scientists to translate ideas into scalable engineering solutions and testing frameworks.
Die vollständige Anzeige sehen
Aufgaben, Anforderungen, Kompetenzen und Vorteile — mit Ihrem kostenlosen Konto.
Bereits ein Konto? Anmelden
Ähnliche Stellen
Weitere Positionen, die passen könnten.
HomeofficePartial
StadtVancouver, Kanada