Senior Backend Engineer - Distributed Systems for ML Inference
仕事内容
Role 1: Design, develop, and deploy scalable, production-grade backend services and distributed systems to support large-scale model inference.
Role 2: Contribute to the technical roadmap of the inference platform, focusing on low-latency, high-throughput services.
Role 3: Ensure reliability and efficiency in production through robust monitoring and observability (Prometheus, Grafana).
Role 4: Collaborate cross-functionally with data science, product, and engineering teams to align platform capabilities with strategic goals.
Role 5: Manage and optimize cloud infrastructure on GCP and orchestrate workloads with Kubernetes.
Role 6: Promote DevOps and SRE best practices across development, testing, deployment, and monitoring; experience with ML inference servers such as NVIDIA Triton is a plus.
Role 2: Contribute to the technical roadmap of the inference platform, focusing on low-latency, high-throughput services.
Role 3: Ensure reliability and efficiency in production through robust monitoring and observability (Prometheus, Grafana).
Role 4: Collaborate cross-functionally with data science, product, and engineering teams to align platform capabilities with strategic goals.
Role 5: Manage and optimize cloud infrastructure on GCP and orchestrate workloads with Kubernetes.
Role 6: Promote DevOps and SRE best practices across development, testing, deployment, and monitoring; experience with ML inference servers such as NVIDIA Triton is a plus.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログイン類似の求人
あなたに合いそうな他の職種。
リモートワークfull
勤務地Remote, Canada