Senior Machine Learning Engineer, On-Device Optimization
Ist diese Stelle etwas für Sie?
Lebenslauf erstellen Erstellen Sie Ihren Lebenslauf und entdecken Sie Ihre Übereinstimmung mit dieser Stelle — und mit allen anderen.
Die Stelle
This role focuses on advancing on-device AI optimization by applying model compression and efficient representation techniques.
You will research and implement quantization, pruning, low-rank factorization, and distillation to reduce model size and improve speed.
Develop deployment methods for state-of-the-art transformer and vision models on-device under hardware constraints.
Lead hardware-aware training strategies to optimize latency, throughput, and memory usage while maintaining accuracy.
Collaborate with software engineers to integrate optimized models into AI companion apps and evaluate cross-framework performance.
Communicate findings through rigorous benchmarking and contribute to open-source ML optimization efforts when possible.
You will research and implement quantization, pruning, low-rank factorization, and distillation to reduce model size and improve speed.
Develop deployment methods for state-of-the-art transformer and vision models on-device under hardware constraints.
Lead hardware-aware training strategies to optimize latency, throughput, and memory usage while maintaining accuracy.
Collaborate with software engineers to integrate optimized models into AI companion apps and evaluate cross-framework performance.
Communicate findings through rigorous benchmarking and contribute to open-source ML optimization efforts when possible.
Die vollständige Anzeige sehen
Aufgaben, Anforderungen, Kompetenzen und Vorteile — mit Ihrem kostenlosen Konto.
Bereits ein Konto? Anmelden
Ähnliche Stellen
Weitere Positionen, die passen könnten.
Homeofficepartial
StadtMultiple locations (US), US