Large Machine Learning Model Optimization Engineer
职位介绍
Lead on-device optimization of large language models and diffusion models to deliver real-time, low-latency experiences.
Drive model compression strategies including quantization, pruning, and distillation for on-device deployment.
Implement and optimize hardware-aware pipelines and ML compilers for efficient inference.
Collaborate with hardware, software, and ML teams to align hardware-software co-design and improve on-device experiences.
Contribute to research publications and share findings with cross-functional teams.
Requirements include Python software engineering, experience with large-scale ML models, and strong collaboration and communication skills.
Drive model compression strategies including quantization, pruning, and distillation for on-device deployment.
Implement and optimize hardware-aware pipelines and ML compilers for efficient inference.
Collaborate with hardware, software, and ML teams to align hardware-software co-design and improve on-device experiences.
Contribute to research publications and share findings with cross-functional teams.
Requirements include Python software engineering, experience with large-scale ML models, and strong collaboration and communication skills.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录相似职位
其他可能适合您的职位。