Large Machine Learning Model Optimization Engineer
Questa offerta fa per te?
Crea il mio CV Crea il tuo CV e scopri la tua percentuale di corrispondenza con questa posizione — e con tutte le altre.
La posizione
Lead on-device optimization of large language models and diffusion models to deliver real-time, low-latency experiences.
Drive model compression strategies including quantization, pruning, and distillation for on-device deployment.
Implement and optimize hardware-aware pipelines and ML compilers for efficient inference.
Collaborate with hardware, software, and ML teams to align hardware-software co-design and improve on-device experiences.
Contribute to research publications and share findings with cross-functional teams.
Requirements include Python software engineering, experience with large-scale ML models, and strong collaboration and communication skills.
Drive model compression strategies including quantization, pruning, and distillation for on-device deployment.
Implement and optimize hardware-aware pipelines and ML compilers for efficient inference.
Collaborate with hardware, software, and ML teams to align hardware-software co-design and improve on-device experiences.
Contribute to research publications and share findings with cross-functional teams.
Requirements include Python software engineering, experience with large-scale ML models, and strong collaboration and communication skills.
Vedi l'annuncio completo
Mansioni, profilo, competenze e vantaggi — crea il tuo account gratuito.
o
Hai già un account?
AccediOfferte simili
Altre posizioni che potrebbero interessarti.