Large Machine Learning Model Optimization Engineer
Cette offre est-elle pour vous ?
Créer mon CV Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.
Le poste
Lead on-device optimization of large language models and diffusion models to deliver real-time, low-latency experiences.
Drive model compression strategies including quantization, pruning, and distillation for on-device deployment.
Implement and optimize hardware-aware pipelines and ML compilers for efficient inference.
Collaborate with hardware, software, and ML teams to align hardware-software co-design and improve on-device experiences.
Contribute to research publications and share findings with cross-functional teams.
Requirements include Python software engineering, experience with large-scale ML models, and strong collaboration and communication skills.
Drive model compression strategies including quantization, pruning, and distillation for on-device deployment.
Implement and optimize hardware-aware pipelines and ML compilers for efficient inference.
Collaborate with hardware, software, and ML teams to align hardware-software co-design and improve on-device experiences.
Contribute to research publications and share findings with cross-functional teams.
Requirements include Python software engineering, experience with large-scale ML models, and strong collaboration and communication skills.
Voir l'offre en entier
Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.
ou
Déjà un compte ?
Se connecterOffres similaires
D'autres postes qui pourraient vous convenir.