AI Models GPU deployment software Engineer
仕事内容
Develop and optimize AI models (LLMs, Vision, MultiModal) on GPUs to boost training and inference performance.
Enhance AI frameworks like PyTorch and TensorFlow in upstream repositories and collaborate with framework maintainers.
Optimize GPU kernels and AI operators (GEMM, Attention) and drive performance on scale-up (multi-GPU) and scale-out (multi-node) systems.
Work in a Linux environment using C++ and Python, applying software engineering best practices, debugging, testing, and performance analysis.
Partner with internal GPU library teams and contribute to open-source initiatives to ensure upstream integration.
Combine independent ownership with collaborative teamwork to deliver robust, scalable, and maintainable code.
Enhance AI frameworks like PyTorch and TensorFlow in upstream repositories and collaborate with framework maintainers.
Optimize GPU kernels and AI operators (GEMM, Attention) and drive performance on scale-up (multi-GPU) and scale-out (multi-node) systems.
Work in a Linux environment using C++ and Python, applying software engineering best practices, debugging, testing, and performance analysis.
Partner with internal GPU library teams and contribute to open-source initiatives to ensure upstream integration.
Combine independent ownership with collaborative teamwork to deliver robust, scalable, and maintainable code.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログイン類似の求人
あなたに合いそうな他の職種。
リモートワークfull
勤務地Remote, United States