AI Models GPU deployment software Engineer
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Develop and optimize AI models (LLMs, Vision, MultiModal) on GPUs to boost training and inference performance.
Enhance AI frameworks like PyTorch and TensorFlow in upstream repositories and collaborate with framework maintainers.
Optimize GPU kernels and AI operators (GEMM, Attention) and drive performance on scale-up (multi-GPU) and scale-out (multi-node) systems.
Work in a Linux environment using C++ and Python, applying software engineering best practices, debugging, testing, and performance analysis.
Partner with internal GPU library teams and contribute to open-source initiatives to ensure upstream integration.
Combine independent ownership with collaborative teamwork to deliver robust, scalable, and maintainable code.
Enhance AI frameworks like PyTorch and TensorFlow in upstream repositories and collaborate with framework maintainers.
Optimize GPU kernels and AI operators (GEMM, Attention) and drive performance on scale-up (multi-GPU) and scale-out (multi-node) systems.
Work in a Linux environment using C++ and Python, applying software engineering best practices, debugging, testing, and performance analysis.
Partner with internal GPU library teams and contribute to open-source initiatives to ensure upstream integration.
Combine independent ownership with collaborative teamwork to deliver robust, scalable, and maintainable code.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inSimilar openings
Other roles that could suit you.
Remote workfull
CityRemote, United States