Senior Research Scientist - Multimodal Agents
Cette offre est-elle pour vous ?
Créer mon CV Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.
Suivez vos candidatures sur mobile L'application Whileresume, gratuite, sur iPhone et Android.
Le poste
Lead research and development of agentic systems spanning planning, tool use, memory, and retrieval for real tasks in design, vision, and language.
Design and optimize reward models and learning loops (RLHF/RLAIF, DPO/IPO) and drive policy optimization across distributed training pipelines.
Build robust evaluation suites, simulate failure modes, and ensure safety, reliability, and scalability of multi-step reasoning.
Advance post-training approaches and MoE architectures while collaborating in a cross-functional environment.
Mentor teammates, share findings, and translate research into high-quality, ship-ready product features.
Partner with safety, product, design, and platform teams to align research with business goals while maintaining rigor and reproducibility.
Design and optimize reward models and learning loops (RLHF/RLAIF, DPO/IPO) and drive policy optimization across distributed training pipelines.
Build robust evaluation suites, simulate failure modes, and ensure safety, reliability, and scalability of multi-step reasoning.
Advance post-training approaches and MoE architectures while collaborating in a cross-functional environment.
Mentor teammates, share findings, and translate research into high-quality, ship-ready product features.
Partner with safety, product, design, and platform teams to align research with business goals while maintaining rigor and reproducibility.
Voir l'offre en entier
Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.
ou
Déjà un compte ?
Se connecterCes offres pourraient vous intéresser
Aucune offre vraiment proche pour l'instant — voici les plus récentes.