Research Engineer, AI Safety & Alignment
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
You will join the Safety team to advance AI safety and alignment for large language models.
Develop and implement evaluation methodologies and metrics to measure safety and alignment.
Research state-of-the-art techniques for model alignment, value learning, and interpretability.
Perform adversarial testing and bias/toxicity mitigation using RLHF and fine-tuning.
Collaborate with engineering and product teams to translate research into scalable safety solutions.
Stay current with the latest AI safety research and contribute to the academic community through publications and talks.
Develop and implement evaluation methodologies and metrics to measure safety and alignment.
Research state-of-the-art techniques for model alignment, value learning, and interpretability.
Perform adversarial testing and bias/toxicity mitigation using RLHF and fine-tuning.
Collaborate with engineering and product teams to translate research into scalable safety solutions.
Stay current with the latest AI safety research and contribute to the academic community through publications and talks.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inSimilar openings
Other roles that could suit you.
? Senior Security Research Engineer ? Research Engineer / Scientist, Frontier Red Team (Cyber) ? Research Engineer, Frontier Red Team (Hardware Lead) ? Research Engineer, Frontier Red Team (Autonomy) ? Research Engineer/Scientist - Generative UI, Consumer Products ? Senior Research Engineer/Scientist - Edge, Consumer Products
Remote workPartial
CityRedwood City, US