ML Engineer - Inference Serving
职位介绍
As an ML Engineer - Inference Serving, you will ship new model architectures by integrating them into our inference engine. Collaborate across research, engineering, and infrastructure to optimize model efficiency and deployment at scale. Build internal tooling to measure, profile, and track the lifetime of inference jobs and workflows. Automate, test, and maintain inference services to ensure maximum uptime and reliability. Design deployment workflows and sophisticated scheduling to optimally leverage GPU resources while meeting SLOs. Develop CI/CD pipelines for processing model checkpoints, platform components, and internal SDKs.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
已有账户? 登录
远程办公no
城市Palo Alto, US