Senior Backend Engineer - Distributed Systems for ML Inference
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Role 1: Design, develop, and deploy scalable, production-grade backend services and distributed systems to support large-scale model inference.
Role 2: Contribute to the technical roadmap of the inference platform, focusing on low-latency, high-throughput services.
Role 3: Ensure reliability and efficiency in production through robust monitoring and observability (Prometheus, Grafana).
Role 4: Collaborate cross-functionally with data science, product, and engineering teams to align platform capabilities with strategic goals.
Role 5: Manage and optimize cloud infrastructure on GCP and orchestrate workloads with Kubernetes.
Role 6: Promote DevOps and SRE best practices across development, testing, deployment, and monitoring; experience with ML inference servers such as NVIDIA Triton is a plus.
Role 2: Contribute to the technical roadmap of the inference platform, focusing on low-latency, high-throughput services.
Role 3: Ensure reliability and efficiency in production through robust monitoring and observability (Prometheus, Grafana).
Role 4: Collaborate cross-functionally with data science, product, and engineering teams to align platform capabilities with strategic goals.
Role 5: Manage and optimize cloud infrastructure on GCP and orchestrate workloads with Kubernetes.
Role 6: Promote DevOps and SRE best practices across development, testing, deployment, and monitoring; experience with ML inference servers such as NVIDIA Triton is a plus.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inSimilar openings
Other roles that could suit you.
Remote workfull
CityRemote, Canada