Staff Machine Learning Performance Engineer, Inference Optimisation
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Lead high-impact projects to optimize ML inference on edge accelerators and GPUs for low-cost, low-power devices.
Advance transformer-based model deployment through ML compiler and kernel optimizations.
Develop cross-platform solutions targeting Nvidia Thor/Orin, Qualcomm, and other SoCs.
Build technical roadmaps and execute with multi-team collaboration across model developers and engineers.
Mentor and guide a growing engineering team, shaping architecture, best practices, and performance benchmarks.
Deliver measurable gains in latency, throughput, and energy efficiency via rigorous testing and benchmarking.
Advance transformer-based model deployment through ML compiler and kernel optimizations.
Develop cross-platform solutions targeting Nvidia Thor/Orin, Qualcomm, and other SoCs.
Build technical roadmaps and execute with multi-team collaboration across model developers and engineers.
Mentor and guide a growing engineering team, shaping architecture, best practices, and performance benchmarks.
Deliver measurable gains in latency, throughput, and energy efficiency via rigorous testing and benchmarking.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inSimilar openings
Other roles that could suit you.