Senior Research Scientist, Reward Models
仕事内容
Lead research on novel reward model architectures and RLHF training approaches for large language models.
Develop and evaluate LLM-based grading and evaluation methods, including rubric-driven approaches that improve consistency and interpretability.
Research techniques to detect, characterize, and mitigate reward hacking and specification gaming.
Design experiments to understand reward model generalization, robustness, and failure modes.
Collaborate with the Finetuning team to translate research insights into improvements for production training pipelines.
Contribute to research publications, blog posts, and internal documentation; mentor other researchers and help build institutional knowledge around reward modeling.
Develop and evaluate LLM-based grading and evaluation methods, including rubric-driven approaches that improve consistency and interpretability.
Research techniques to detect, characterize, and mitigate reward hacking and specification gaming.
Design experiments to understand reward model generalization, robustness, and failure modes.
Collaborate with the Finetuning team to translate research insights into improvements for production training pipelines.
Contribute to research publications, blog posts, and internal documentation; mentor other researchers and help build institutional knowledge around reward modeling.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログイン類似の求人
あなたに合いそうな他の職種。
リモートワークfull
勤務地Remote (Canada), Canada