Research Engineer, Production Model Post-Training - London
职位介绍
Implement and optimize post-training techniques at scale on frontier language models to improve production safety, alignment, and performance.
Conduct research to develop post-training recipes, including Constitutional AI and RLHF, that directly enhance model quality.
Design, build, and run robust pipelines for fine-tuning and evaluating models across diverse metrics.
Develop tools to measure, monitor, and debug model training and deployment in large-scale distributed systems.
Collaborate with research and engineering teams to translate cutting-edge techniques into production-ready implementations and establish best practices for reliability.
Balance research exploration with engineering rigor, maintain clarity under time pressure, and contribute to safe, responsible AI deployment.
Conduct research to develop post-training recipes, including Constitutional AI and RLHF, that directly enhance model quality.
Design, build, and run robust pipelines for fine-tuning and evaluating models across diverse metrics.
Develop tools to measure, monitor, and debug model training and deployment in large-scale distributed systems.
Collaborate with research and engineering teams to translate cutting-edge techniques into production-ready implementations and establish best practices for reliability.
Balance research exploration with engineering rigor, maintain clarity under time pressure, and contribute to safe, responsible AI deployment.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录相似职位
其他可能适合您的职位。
远程办公partial
城市London, 英国