Principal Machine Learning Researcher, On-Device Optimization
仕事内容
As a Principal ML Researcher focused on on-device optimization, you will bridge advanced research and product-ready deployment for edge AI systems.
You will lead research and implementation of model compression techniques such as quantization, pruning, distillation, and low-rank factorization.
You will develop methods to run state-of-the-art transformer and vision models on-device under hardware constraints.
You will drive hardware-aware training strategies to optimize latency, throughput, and memory usage, and collaborate with software engineers to integrate models into applications.
You will evaluate frameworks and quantization strategies (e.g., AWQ, GPTQ, SmoothQuant) and benchmark performance.
The role requires a PhD with related experience, expertise in edge ML, and strong collaboration and communication skills.
You will lead research and implementation of model compression techniques such as quantization, pruning, distillation, and low-rank factorization.
You will develop methods to run state-of-the-art transformer and vision models on-device under hardware constraints.
You will drive hardware-aware training strategies to optimize latency, throughput, and memory usage, and collaborate with software engineers to integrate models into applications.
You will evaluate frameworks and quantization strategies (e.g., AWQ, GPTQ, SmoothQuant) and benchmark performance.
The role requires a PhD with related experience, expertise in edge ML, and strong collaboration and communication skills.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
すでにアカウントをお持ちですか? ログイン
類似の求人
あなたに合いそうな他の職種。
リモートワークno
勤務地Multiple locations, United States