Skip to content
Whileresume
Leaderboard Build my CV Hire Log in

Staff Machine Learning Engineer, ML Performance & Optimization

Is this job for you?

Build your CV and see how well you match this role — and every other one.

Build my CV
Track your applications on mobile The free Whileresume app, on iPhone and Android.

The role

Take ownership as a Staff ML Engineer focused on ML performance and optimization across multi-platform deployments.
You will optimize neural architectures and systems for high performance on GPU/TPU hardware, including onboard and simulation platforms.
Develop post-training techniques like quantization and kernel-level optimizations to reduce latency and memory footprint for real-time constraints.
Experiment with new architectures (sparse models) and decoding strategies (speculative decoding) to boost inference speed.
Enhance training efficiency for large models and fine-tuning in data-heavy pipelines, while collaborating with ML infra, hardware, and research teams.
This hybrid role emphasizes hands-on optimization and cross-team collaboration to scale production-grade models.

See the full job post

Responsibilities, requirements, skills and benefits — create your free account.

At least 6 characters. The longer, the safer.
or

Already have an account?

You might also like these jobs

No closely matching jobs yet — here are the most recent ones.

Your location

Jobs and companies will be filtered on this country.

Suggested

All countries 65