Research Engineer, AI Safety & Alignment
职位介绍
You will join the Safety team to advance AI safety and alignment for large language models.
Develop and implement evaluation methodologies and metrics to measure safety and alignment.
Research state-of-the-art techniques for model alignment, value learning, and interpretability.
Perform adversarial testing and bias/toxicity mitigation using RLHF and fine-tuning.
Collaborate with engineering and product teams to translate research into scalable safety solutions.
Stay current with the latest AI safety research and contribute to the academic community through publications and talks.
Develop and implement evaluation methodologies and metrics to measure safety and alignment.
Research state-of-the-art techniques for model alignment, value learning, and interpretability.
Perform adversarial testing and bias/toxicity mitigation using RLHF and fine-tuning.
Collaborate with engineering and product teams to translate research into scalable safety solutions.
Stay current with the latest AI safety research and contribute to the academic community through publications and talks.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录相似职位
其他可能适合您的职位。
? Senior Security Research Engineer ? Research Engineer / Scientist, Frontier Red Team (Cyber) ? Research Engineer, Frontier Red Team (Hardware Lead) ? Research Engineer, Frontier Red Team (Autonomy) ? Research Engineer/Scientist - Generative UI, Consumer Products ? Senior Research Engineer/Scientist - Edge, Consumer Products
远程办公Partial
城市Redwood City, US