Skip to content

Staff Machine Learning Performance Engineer, Inference Optimisation

Is this job for you?

Build your CV and see how well you match this role — and every other one.

Build my CV

The role

Lead high-impact projects to optimize ML inference on edge accelerators and GPUs for low-cost, low-power devices.
Advance transformer-based model deployment through ML compiler and kernel optimizations.
Develop cross-platform solutions targeting Nvidia Thor/Orin, Qualcomm, and other SoCs.
Build technical roadmaps and execute with multi-team collaboration across model developers and engineers.
Mentor and guide a growing engineering team, shaping architecture, best practices, and performance benchmarks.
Deliver measurable gains in latency, throughput, and energy efficiency via rigorous testing and benchmarking.

See the full job post

Responsibilities, requirements, skills and benefits — create your free account.

At least 6 characters. The longer, the safer.
or

Already have an account?

Similar openings

Other roles that could suit you.

See all →

Your location

Jobs and companies will be filtered on this country.

Suggested

All countries 65