Senior Backend Engineer - Distributed Systems for ML Inference
职位介绍
Role 1: Design, develop, and deploy scalable, production-grade backend services and distributed systems to support large-scale model inference.
Role 2: Contribute to the technical roadmap of the inference platform, focusing on low-latency, high-throughput services.
Role 3: Ensure reliability and efficiency in production through robust monitoring and observability (Prometheus, Grafana).
Role 4: Collaborate cross-functionally with data science, product, and engineering teams to align platform capabilities with strategic goals.
Role 5: Manage and optimize cloud infrastructure on GCP and orchestrate workloads with Kubernetes.
Role 6: Promote DevOps and SRE best practices across development, testing, deployment, and monitoring; experience with ML inference servers such as NVIDIA Triton is a plus.
Role 2: Contribute to the technical roadmap of the inference platform, focusing on low-latency, high-throughput services.
Role 3: Ensure reliability and efficiency in production through robust monitoring and observability (Prometheus, Grafana).
Role 4: Collaborate cross-functionally with data science, product, and engineering teams to align platform capabilities with strategic goals.
Role 5: Manage and optimize cloud infrastructure on GCP and orchestrate workloads with Kubernetes.
Role 6: Promote DevOps and SRE best practices across development, testing, deployment, and monitoring; experience with ML inference servers such as NVIDIA Triton is a plus.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录相似职位
其他可能适合您的职位。
远程办公full
城市Remote, Canada