Research Engineer, AI Safety & Alignment
仕事内容
You will join the Safety team to advance AI safety and alignment for large language models.
Develop and implement evaluation methodologies and metrics to measure safety and alignment.
Research state-of-the-art techniques for model alignment, value learning, and interpretability.
Perform adversarial testing and bias/toxicity mitigation using RLHF and fine-tuning.
Collaborate with engineering and product teams to translate research into scalable safety solutions.
Stay current with the latest AI safety research and contribute to the academic community through publications and talks.
Develop and implement evaluation methodologies and metrics to measure safety and alignment.
Research state-of-the-art techniques for model alignment, value learning, and interpretability.
Perform adversarial testing and bias/toxicity mitigation using RLHF and fine-tuning.
Collaborate with engineering and product teams to translate research into scalable safety solutions.
Stay current with the latest AI safety research and contribute to the academic community through publications and talks.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログイン類似の求人
あなたに合いそうな他の職種。
? Senior Security Research Engineer ? Research Engineer / Scientist, Frontier Red Team (Cyber) ? Research Engineer, Frontier Red Team (Hardware Lead) ? Research Engineer, Frontier Red Team (Autonomy) ? Research Engineer/Scientist - Generative UI, Consumer Products ? Senior Research Engineer/Scientist - Edge, Consumer Products
リモートワークPartial
勤務地Redwood City, US