Skip to content

Machine Learning Infrastructure Engineer

Is this job for you?

Build your CV and see how well you match this role — and every other one.

Build my CV

The role

The role focuses on designing, building, and maintaining scalable ML training and serving infrastructure to accelerate research and product development.
You will develop tooling to diagnose cluster issues and hardware failures, monitor deployments, and manage experiments.
A core priority is maximizing GPU allocation and utilization for both training and serving workloads.
Provide infrastructure support to ML teams, troubleshoot performance bottlenecks, and coordinate with platform engineers.
Required hands-on experience with cloud platforms, Kubernetes, and GPU-enabled environments, plus experience with PyTorch, TensorFlow, or JAX.
This role demands 4+ years of ML infra experience and a proactive, collaborative mindset in a fast-paced research setting.

See the full job post

Responsibilities, requirements, skills and benefits — create your free account.

At least 6 characters. The longer, the safer.
or

Already have an account?

Similar openings

Other roles that could suit you.

See all →

Your location

Jobs and companies will be filtered on this country.

Suggested

All countries 65