Staff Machine Learning Performance Engineer, Inference Optimisation
仕事内容
Lead high-impact projects to optimize ML inference on edge accelerators and GPUs for low-cost, low-power devices.
Advance transformer-based model deployment through ML compiler and kernel optimizations.
Develop cross-platform solutions targeting Nvidia Thor/Orin, Qualcomm, and other SoCs.
Build technical roadmaps and execute with multi-team collaboration across model developers and engineers.
Mentor and guide a growing engineering team, shaping architecture, best practices, and performance benchmarks.
Deliver measurable gains in latency, throughput, and energy efficiency via rigorous testing and benchmarking.
Advance transformer-based model deployment through ML compiler and kernel optimizations.
Develop cross-platform solutions targeting Nvidia Thor/Orin, Qualcomm, and other SoCs.
Build technical roadmaps and execute with multi-team collaboration across model developers and engineers.
Mentor and guide a growing engineering team, shaping architecture, best practices, and performance benchmarks.
Deliver measurable gains in latency, throughput, and energy efficiency via rigorous testing and benchmarking.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログイン類似の求人
あなたに合いそうな他の職種。