Research Engineer, AI Safety & Alignment
Is deze vacature iets voor u?
Mijn cv maken Maak uw cv en ontdek uw matchpercentage met deze functie — en met alle andere.
Volg uw sollicitaties op mobiel De gratis Whileresume-app, op iPhone en Android.
De functie
Develop and implement novel evaluation methodologies and metrics to assess safety and alignment of large language models.
Research and build techniques for model alignment, value learning, and interpretability.
Conduct adversarial testing to proactively uncover vulnerabilities and failure modes.
Analyze and mitigate biases, toxicity, and other harmful behaviors through RLHF and fine-tuning.
Collaborate with engineering and product teams to translate safety research into scalable, production-ready solutions.
Stay current with AI safety advances and contribute to the academic community through publications and talks.
Research and build techniques for model alignment, value learning, and interpretability.
Conduct adversarial testing to proactively uncover vulnerabilities and failure modes.
Analyze and mitigate biases, toxicity, and other harmful behaviors through RLHF and fine-tuning.
Collaborate with engineering and product teams to translate safety research into scalable, production-ready solutions.
Stay current with AI safety advances and contribute to the academic community through publications and talks.
Bekijk de volledige vacature
Taken, profiel, vaardigheden en voordelen — maak gratis een account aan.
of
Al een account?
InloggenDeze vacatures zijn misschien iets voor u
Nog geen echt vergelijkbare vacatures — hier zijn de nieuwste.