Research Engineer, AI Safety & Alignment
Is deze vacature iets voor u?
Mijn cv maken Maak uw cv en ontdek uw matchpercentage met deze functie — en met alle andere.
De functie
You will join the Safety team to advance AI safety and alignment for large language models.
Develop and implement evaluation methodologies and metrics to measure safety and alignment.
Research state-of-the-art techniques for model alignment, value learning, and interpretability.
Perform adversarial testing and bias/toxicity mitigation using RLHF and fine-tuning.
Collaborate with engineering and product teams to translate research into scalable safety solutions.
Stay current with the latest AI safety research and contribute to the academic community through publications and talks.
Develop and implement evaluation methodologies and metrics to measure safety and alignment.
Research state-of-the-art techniques for model alignment, value learning, and interpretability.
Perform adversarial testing and bias/toxicity mitigation using RLHF and fine-tuning.
Collaborate with engineering and product teams to translate research into scalable safety solutions.
Stay current with the latest AI safety research and contribute to the academic community through publications and talks.
Bekijk de volledige vacature
Taken, profiel, vaardigheden en voordelen — maak gratis een account aan.
of
Al een account?
InloggenVergelijkbare vacatures
Andere functies die kunnen passen.
? Senior Security Research Engineer ? Research Engineer / Scientist, Frontier Red Team (Cyber) ? Research Engineer, Frontier Red Team (Hardware Lead) ? Research Engineer, Frontier Red Team (Autonomy) ? Research Engineer/Scientist - Generative UI, Consumer Products ? Senior Research Engineer/Scientist - Edge, Consumer Products
ThuiswerkenPartial
StadRedwood City, US