Staff Software Engineer, ML Frameworks & Efficiency
在手机上跟进您的申请 Whileresume 免费应用,支持 iPhone 和 Android。
职位介绍
Optimize distributed ML systems for high performance on TPU and GPU clusters, applying SPMD, MPMD, and FSDP to scale model training.
Improve accelerator FLOPS efficiency by refining compiler optimizations (XLA), authoring low-level kernels (Pallas, Triton), and enabling low-precision computation.
Develop new neural model architectures (e.g., sparse architectures) and decoding strategies (e.g., speculative decoding) for improved training and inference on modern hardware.
Evaluate and integrate open-source and state-of-the-art technologies to enhance the performance and scalability of ML workloads.
Promote best practices for distributed systems architecture and contribute to technical leadership within the team.
Qualifications include a strong CS/math background, experience with ML frameworks (TensorFlow, JAX, XLA), Python and C++, and proficiency with profiling tools.
Improve accelerator FLOPS efficiency by refining compiler optimizations (XLA), authoring low-level kernels (Pallas, Triton), and enabling low-precision computation.
Develop new neural model architectures (e.g., sparse architectures) and decoding strategies (e.g., speculative decoding) for improved training and inference on modern hardware.
Evaluate and integrate open-source and state-of-the-art technologies to enhance the performance and scalability of ML workloads.
Promote best practices for distributed systems architecture and contribute to technical leadership within the team.
Qualifications include a strong CS/math background, experience with ML frameworks (TensorFlow, JAX, XLA), Python and C++, and proficiency with profiling tools.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录您可能也感兴趣的职位
暂无高度相似的职位 — 以下是最新职位。