ML Engineer - Inference Serving
仕事内容
As an ML Engineer - Inference Serving, you will ship new model architectures by integrating them into our inference engine. Collaborate across research, engineering, and infrastructure to optimize model efficiency and deployment at scale. Build internal tooling to measure, profile, and track the lifetime of inference jobs and workflows. Automate, test, and maintain inference services to ensure maximum uptime and reliability. Design deployment workflows and sophisticated scheduling to optimally leverage GPU resources while meeting SLOs. Develop CI/CD pipelines for processing model checkpoints, platform components, and internal SDKs.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
すでにアカウントをお持ちですか? ログイン
リモートワークno
勤務地Palo Alto, US