Research Engineer, AI Safety & Alignment
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
Track your applications on mobile The free Whileresume app, on iPhone and Android.
The role
You will join the Safety team to advance AI safety and alignment for large language models.
Develop and implement evaluation methodologies and metrics to measure safety and alignment.
Research state-of-the-art techniques for model alignment, value learning, and interpretability.
Perform adversarial testing and bias/toxicity mitigation using RLHF and fine-tuning.
Collaborate with engineering and product teams to translate research into scalable safety solutions.
Stay current with the latest AI safety research and contribute to the academic community through publications and talks.
Develop and implement evaluation methodologies and metrics to measure safety and alignment.
Research state-of-the-art techniques for model alignment, value learning, and interpretability.
Perform adversarial testing and bias/toxicity mitigation using RLHF and fine-tuning.
Collaborate with engineering and product teams to translate research into scalable safety solutions.
Stay current with the latest AI safety research and contribute to the academic community through publications and talks.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inYou might also like these jobs
No closely matching jobs yet — here are the most recent ones.