Member of Technical Staff, Frontiers of Deep Learning Scaling
职位介绍
Lead research and engineering efforts to identify and validate scalable paradigms for efficient compute and useful data in next-token prediction.
Design, implement, and run large-scale, long-duration training workflows with robust stability across hundreds of millions of GPU hours.
Build end-to-end pipelines: data preparation, evaluation, experiment design, result analysis, and iterative redesign for faster learning.
Explore novel architectures, learning paradigms (e.g., continual learning, self-improvement), and unified multi-modal models as potential scaling paths.
Collaborate across teams, communicate results concisely, and drive fast, principled experimentation and decision making.
Work with Python, JAX, PyTorch, and Rust to implement ideas, test hypotheses, and deliver measurable improvements.
Design, implement, and run large-scale, long-duration training workflows with robust stability across hundreds of millions of GPU hours.
Build end-to-end pipelines: data preparation, evaluation, experiment design, result analysis, and iterative redesign for faster learning.
Explore novel architectures, learning paradigms (e.g., continual learning, self-improvement), and unified multi-modal models as potential scaling paths.
Collaborate across teams, communicate results concisely, and drive fast, principled experimentation and decision making.
Work with Python, JAX, PyTorch, and Rust to implement ideas, test hypotheses, and deliver measurable improvements.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录远程办公Partial
城市Palo Alto, United States