Staff Software Engineer, AI Runtime Systems
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Lead the design and evolution of AI inference runtimes and core platform services to support scalable, secure ML workloads.
Architect multi-model serving with autoscaling and GPU/TPU utilization while optimizing latency and throughput.
Collaborate with ML and data science teams to deploy, monitor, and operationalize production models.
Shape authentication, security, and notification services, ensuring reliable and compliant production operations.
Ensure seamless CI/CD integration, observability, benchmarking, and cross-team design alignment.
Mentor engineers and champion engineering excellence, setting standards across platform and ML services.
Architect multi-model serving with autoscaling and GPU/TPU utilization while optimizing latency and throughput.
Collaborate with ML and data science teams to deploy, monitor, and operationalize production models.
Shape authentication, security, and notification services, ensuring reliable and compliant production operations.
Ensure seamless CI/CD integration, observability, benchmarking, and cross-team design alignment.
Mentor engineers and champion engineering excellence, setting standards across platform and ML services.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inSimilar openings
Other roles that could suit you.