Research Engineer, AI Safety & Alignment
Cette offre est-elle pour vous ?
Créer mon CV Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.
Suivez vos candidatures sur mobile L'application Whileresume, gratuite, sur iPhone et Android.
Le poste
Develop and implement novel evaluation methodologies and metrics to assess safety and alignment of large language models.
Research and build techniques for model alignment, value learning, and interpretability.
Conduct adversarial testing to proactively uncover vulnerabilities and failure modes.
Analyze and mitigate biases, toxicity, and other harmful behaviors through RLHF and fine-tuning.
Collaborate with engineering and product teams to translate safety research into scalable, production-ready solutions.
Stay current with AI safety advances and contribute to the academic community through publications and talks.
Research and build techniques for model alignment, value learning, and interpretability.
Conduct adversarial testing to proactively uncover vulnerabilities and failure modes.
Analyze and mitigate biases, toxicity, and other harmful behaviors through RLHF and fine-tuning.
Collaborate with engineering and product teams to translate safety research into scalable, production-ready solutions.
Stay current with AI safety advances and contribute to the academic community through publications and talks.
Voir l'offre en entier
Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.
ou
Déjà un compte ?
Se connecterCes offres pourraient vous intéresser
Aucune offre vraiment proche pour l'instant — voici les plus récentes.