Staff Machine Learning Performance Engineer, Inference Optimisation
职位介绍
Lead high-impact projects to optimize ML inference on edge accelerators and GPUs for low-cost, low-power devices.
Advance transformer-based model deployment through ML compiler and kernel optimizations.
Develop cross-platform solutions targeting Nvidia Thor/Orin, Qualcomm, and other SoCs.
Build technical roadmaps and execute with multi-team collaboration across model developers and engineers.
Mentor and guide a growing engineering team, shaping architecture, best practices, and performance benchmarks.
Deliver measurable gains in latency, throughput, and energy efficiency via rigorous testing and benchmarking.
Advance transformer-based model deployment through ML compiler and kernel optimizations.
Develop cross-platform solutions targeting Nvidia Thor/Orin, Qualcomm, and other SoCs.
Build technical roadmaps and execute with multi-team collaboration across model developers and engineers.
Mentor and guide a growing engineering team, shaping architecture, best practices, and performance benchmarks.
Deliver measurable gains in latency, throughput, and energy efficiency via rigorous testing and benchmarking.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录相似职位
其他可能适合您的职位。