Principal Machine Learning Researcher, On-Device Optimization
Cette offre est-elle pour vous ?
Créer mon CV Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.
Le poste
As a Principal ML Researcher focused on on-device optimization, you will bridge advanced research and product-ready deployment for edge AI systems.
You will lead research and implementation of model compression techniques such as quantization, pruning, distillation, and low-rank factorization.
You will develop methods to run state-of-the-art transformer and vision models on-device under hardware constraints.
You will drive hardware-aware training strategies to optimize latency, throughput, and memory usage, and collaborate with software engineers to integrate models into applications.
You will evaluate frameworks and quantization strategies (e.g., AWQ, GPTQ, SmoothQuant) and benchmark performance.
The role requires a PhD with related experience, expertise in edge ML, and strong collaboration and communication skills.
You will lead research and implementation of model compression techniques such as quantization, pruning, distillation, and low-rank factorization.
You will develop methods to run state-of-the-art transformer and vision models on-device under hardware constraints.
You will drive hardware-aware training strategies to optimize latency, throughput, and memory usage, and collaborate with software engineers to integrate models into applications.
You will evaluate frameworks and quantization strategies (e.g., AWQ, GPTQ, SmoothQuant) and benchmark performance.
The role requires a PhD with related experience, expertise in edge ML, and strong collaboration and communication skills.
Voir l'offre en entier
Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.
Déjà un compte ? Se connecter
Offres similaires
D'autres postes qui pourraient vous convenir.
Télétravailno
VilleMultiple locations, United States