跳到正文

Staff Software Engineer, ML Frameworks & Efficiency

这个职位适合你吗?

创建简历,即可看到你与这个职位——以及其他所有职位——的匹配度。

创建我的简历

职位介绍

Optimize distributed ML systems for high performance on TPU and GPU clusters, applying SPMD, MPMD, and FSDP to scale model training.
Improve accelerator FLOPS efficiency by refining compiler optimizations (XLA), authoring low-level kernels (Pallas, Triton), and enabling low-precision computation.
Develop new neural model architectures (e.g., sparse architectures) and decoding strategies (e.g., speculative decoding) for improved training and inference on modern hardware.
Evaluate and integrate open-source and state-of-the-art technologies to enhance the performance and scalability of ML workloads.
Promote best practices for distributed systems architecture and contribute to technical leadership within the team.
Qualifications include a strong CS/math background, experience with ML frameworks (TensorFlow, JAX, XLA), Python and C++, and proficiency with profiling tools.

查看完整职位

工作职责、任职要求、技能与福利 — 免费创建账号即可查看。

至少 6 个字符。越长越安全。

已有账户? 登录

相似职位

其他可能适合您的职位。

查看全部 →

您的城市

职位和企业将按该国家筛选。

推荐

所有国家 66