Research Engineer, AI Safety & Alignment
Ist diese Stelle etwas für Sie?
Lebenslauf erstellen Erstellen Sie Ihren Lebenslauf und entdecken Sie Ihre Übereinstimmung mit dieser Stelle — und mit allen anderen.
Bewerbungen unterwegs verfolgen Die kostenlose Whileresume-App für iPhone und Android.
Die Stelle
Develop and implement novel evaluation methodologies and metrics to assess safety and alignment of large language models.
Research and build techniques for model alignment, value learning, and interpretability.
Conduct adversarial testing to proactively uncover vulnerabilities and failure modes.
Analyze and mitigate biases, toxicity, and other harmful behaviors through RLHF and fine-tuning.
Collaborate with engineering and product teams to translate safety research into scalable, production-ready solutions.
Stay current with AI safety advances and contribute to the academic community through publications and talks.
Research and build techniques for model alignment, value learning, and interpretability.
Conduct adversarial testing to proactively uncover vulnerabilities and failure modes.
Analyze and mitigate biases, toxicity, and other harmful behaviors through RLHF and fine-tuning.
Collaborate with engineering and product teams to translate safety research into scalable, production-ready solutions.
Stay current with AI safety advances and contribute to the academic community through publications and talks.
Die vollständige Anzeige sehen
Aufgaben, Anforderungen, Kompetenzen und Vorteile — mit Ihrem kostenlosen Konto.
oder
Bereits ein Konto?
AnmeldenDiese Stellen könnten Sie interessieren
Noch keine wirklich ähnliche Stelle — hier die neuesten.