Large Machine Learning Model Optimization Engineer
Is deze vacature iets voor u?
Mijn cv maken Maak uw cv en ontdek uw matchpercentage met deze functie — en met alle andere.
De functie
Lead on-device optimization of large language models and diffusion models to deliver real-time, low-latency experiences.
Drive model compression strategies including quantization, pruning, and distillation for on-device deployment.
Implement and optimize hardware-aware pipelines and ML compilers for efficient inference.
Collaborate with hardware, software, and ML teams to align hardware-software co-design and improve on-device experiences.
Contribute to research publications and share findings with cross-functional teams.
Requirements include Python software engineering, experience with large-scale ML models, and strong collaboration and communication skills.
Drive model compression strategies including quantization, pruning, and distillation for on-device deployment.
Implement and optimize hardware-aware pipelines and ML compilers for efficient inference.
Collaborate with hardware, software, and ML teams to align hardware-software co-design and improve on-device experiences.
Contribute to research publications and share findings with cross-functional teams.
Requirements include Python software engineering, experience with large-scale ML models, and strong collaboration and communication skills.
Bekijk de volledige vacature
Taken, profiel, vaardigheden en voordelen — maak gratis een account aan.
of
Al een account?
InloggenVergelijkbare vacatures
Andere functies die kunnen passen.