Research Engineer, ML Systems (All Industry Levels)
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Join the ML Systems team to optimize GPU-based AI training and inference at scale.
Develop efficient kernels (Triton, CUDA) and tune performance for models and hardware.
Improve serving with prefix-aware routing and cache-hit optimization for 20K+ QPS.
Train and distill LLMs to reduce latency while maintaining accuracy and engagement.
Build scalable distributed RLHF pipelines and multimodal model training/inference systems.
Collaborate across teams, write clean production-grade code, and contribute to cutting-edge AI solutions.
Develop efficient kernels (Triton, CUDA) and tune performance for models and hardware.
Improve serving with prefix-aware routing and cache-hit optimization for 20K+ QPS.
Train and distill LLMs to reduce latency while maintaining accuracy and engagement.
Build scalable distributed RLHF pipelines and multimodal model training/inference systems.
Collaborate across teams, write clean production-grade code, and contribute to cutting-edge AI solutions.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inSimilar openings
Other roles that could suit you.
? Senior Security Research Engineer ? Research Engineer / Scientist, Frontier Red Team (Cyber) ? Research Engineer, Frontier Red Team (Hardware Lead) ? Research Engineer, Frontier Red Team (Autonomy) ? Research Engineer/Scientist - Generative UI, Consumer Products ? Senior Research Engineer/Scientist - Edge, Consumer Products
Remote workPartial
CityRedwood City, US