Research Engineer, ML Systems (All Industry Levels)
Cette offre est-elle pour vous ?
Créer mon CV Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.
Suivez vos candidatures sur mobile L'application Whileresume, gratuite, sur iPhone et Android.
Le poste
Join the ML Systems team to optimize GPU-based AI training and inference at scale.
Develop efficient kernels (Triton, CUDA) and tune performance for models and hardware.
Improve serving with prefix-aware routing and cache-hit optimization for 20K+ QPS.
Train and distill LLMs to reduce latency while maintaining accuracy and engagement.
Build scalable distributed RLHF pipelines and multimodal model training/inference systems.
Collaborate across teams, write clean production-grade code, and contribute to cutting-edge AI solutions.
Develop efficient kernels (Triton, CUDA) and tune performance for models and hardware.
Improve serving with prefix-aware routing and cache-hit optimization for 20K+ QPS.
Train and distill LLMs to reduce latency while maintaining accuracy and engagement.
Build scalable distributed RLHF pipelines and multimodal model training/inference systems.
Collaborate across teams, write clean production-grade code, and contribute to cutting-edge AI solutions.
Voir l'offre en entier
Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.
ou
Déjà un compte ?
Se connecterCes offres pourraient vous intéresser
Aucune offre vraiment proche pour l'instant — voici les plus récentes.