Skip to content

AI Models GPU deployment software Engineer

Is this job for you?

Build your CV and see how well you match this role — and every other one.

Build my CV

The role

Develop and optimize AI models (LLMs, Vision, MultiModal) on GPUs to boost training and inference performance.
Enhance AI frameworks like PyTorch and TensorFlow in upstream repositories and collaborate with framework maintainers.
Optimize GPU kernels and AI operators (GEMM, Attention) and drive performance on scale-up (multi-GPU) and scale-out (multi-node) systems.
Work in a Linux environment using C++ and Python, applying software engineering best practices, debugging, testing, and performance analysis.
Partner with internal GPU library teams and contribute to open-source initiatives to ensure upstream integration.
Combine independent ownership with collaborative teamwork to deliver robust, scalable, and maintainable code.

See the full job post

Responsibilities, requirements, skills and benefits — create your free account.

At least 6 characters. The longer, the safer.
or

Already have an account?

Similar openings

Other roles that could suit you.

See all →

Your location

Jobs and companies will be filtered on this country.

Suggested

All countries 65