Senior Research Scientist, Reward Models
Is deze vacature iets voor u?
Mijn cv maken Maak uw cv en ontdek uw matchpercentage met deze functie — en met alle andere.
Volg uw sollicitaties op mobiel De gratis Whileresume-app, op iPhone en Android.
De functie
Lead research on novel reward model architectures and RLHF training approaches for large language models.
Develop and evaluate LLM-based grading and evaluation methods, including rubric-driven approaches that improve consistency and interpretability.
Research techniques to detect, characterize, and mitigate reward hacking and specification gaming.
Design experiments to understand reward model generalization, robustness, and failure modes.
Collaborate with the Finetuning team to translate research insights into improvements for production training pipelines.
Contribute to research publications, blog posts, and internal documentation; mentor other researchers and help build institutional knowledge around reward modeling.
Develop and evaluate LLM-based grading and evaluation methods, including rubric-driven approaches that improve consistency and interpretability.
Research techniques to detect, characterize, and mitigate reward hacking and specification gaming.
Design experiments to understand reward model generalization, robustness, and failure modes.
Collaborate with the Finetuning team to translate research insights into improvements for production training pipelines.
Contribute to research publications, blog posts, and internal documentation; mentor other researchers and help build institutional knowledge around reward modeling.
Bekijk de volledige vacature
Taken, profiel, vaardigheden en voordelen — maak gratis een account aan.
of
Al een account?
InloggenDeze vacatures zijn misschien iets voor u
Nog geen echt vergelijkbare vacatures — hier zijn de nieuwste.