Principal Machine Learning Researcher, On-Device Optimization
Is deze vacature iets voor u?
Mijn cv maken Maak uw cv en ontdek uw matchpercentage met deze functie — en met alle andere.
De functie
As a Principal ML Researcher focused on on-device optimization, you will bridge advanced research and product-ready deployment for edge AI systems.
You will lead research and implementation of model compression techniques such as quantization, pruning, distillation, and low-rank factorization.
You will develop methods to run state-of-the-art transformer and vision models on-device under hardware constraints.
You will drive hardware-aware training strategies to optimize latency, throughput, and memory usage, and collaborate with software engineers to integrate models into applications.
You will evaluate frameworks and quantization strategies (e.g., AWQ, GPTQ, SmoothQuant) and benchmark performance.
The role requires a PhD with related experience, expertise in edge ML, and strong collaboration and communication skills.
You will lead research and implementation of model compression techniques such as quantization, pruning, distillation, and low-rank factorization.
You will develop methods to run state-of-the-art transformer and vision models on-device under hardware constraints.
You will drive hardware-aware training strategies to optimize latency, throughput, and memory usage, and collaborate with software engineers to integrate models into applications.
You will evaluate frameworks and quantization strategies (e.g., AWQ, GPTQ, SmoothQuant) and benchmark performance.
The role requires a PhD with related experience, expertise in edge ML, and strong collaboration and communication skills.
Bekijk de volledige vacature
Taken, profiel, vaardigheden en voordelen — maak gratis een account aan.
Al een account? Inloggen
Vergelijkbare vacatures
Andere functies die kunnen passen.
Thuiswerkenno
StadMultiple locations, United States