Research Engineer, ML Systems (All Industry Levels)
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
Track your applications on mobile The free Whileresume app, on iPhone and Android.
The role
Join the ML Systems team to optimize GPU-based AI training and inference at scale.
Develop efficient kernels (Triton, CUDA) and tune performance for models and hardware.
Improve serving with prefix-aware routing and cache-hit optimization for 20K+ QPS.
Train and distill LLMs to reduce latency while maintaining accuracy and engagement.
Build scalable distributed RLHF pipelines and multimodal model training/inference systems.
Collaborate across teams, write clean production-grade code, and contribute to cutting-edge AI solutions.
Develop efficient kernels (Triton, CUDA) and tune performance for models and hardware.
Improve serving with prefix-aware routing and cache-hit optimization for 20K+ QPS.
Train and distill LLMs to reduce latency while maintaining accuracy and engagement.
Build scalable distributed RLHF pipelines and multimodal model training/inference systems.
Collaborate across teams, write clean production-grade code, and contribute to cutting-edge AI solutions.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inYou might also like these jobs
No closely matching jobs yet — here are the most recent ones.