Research Engineer, AI Safety & Alignment
Cette offre est-elle pour vous ?
Créer mon CV Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.
Le poste
Develop and implement novel evaluation methodologies and metrics to assess safety and alignment of large language models.
Research and build techniques for model alignment, value learning, and interpretability.
Conduct adversarial testing to proactively uncover vulnerabilities and failure modes.
Analyze and mitigate biases, toxicity, and other harmful behaviors through RLHF and fine-tuning.
Collaborate with engineering and product teams to translate safety research into scalable, production-ready solutions.
Stay current with AI safety advances and contribute to the academic community through publications and talks.
Research and build techniques for model alignment, value learning, and interpretability.
Conduct adversarial testing to proactively uncover vulnerabilities and failure modes.
Analyze and mitigate biases, toxicity, and other harmful behaviors through RLHF and fine-tuning.
Collaborate with engineering and product teams to translate safety research into scalable, production-ready solutions.
Stay current with AI safety advances and contribute to the academic community through publications and talks.
Voir l'offre en entier
Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.
Déjà un compte ? Se connecter
Offres similaires
D'autres postes qui pourraient vous convenir.
? Senior Security Research Engineer ? Research Engineer / Scientist, Frontier Red Team (Cyber) ? Research Engineer, Frontier Red Team (Hardware Lead) ? Research Engineer, Frontier Red Team (Autonomy) ? Research Engineer/Scientist - Generative UI, Consumer Products ? Senior Research Engineer/Scientist - Edge, Consumer Products
Télétravailno
VilleRedwood City, US