Ir al contenido
Whileresume
Clasificación Crear mi CV Contratar Iniciar sesión

Staff ML Serving Platform Engineer

¿Es esta oferta para usted?

Cree su CV y descubra su porcentaje de coincidencia con este puesto — y con todos los demás.

Crear mi CV
Sigue tus candidaturas desde el móvil La app gratuita de Whileresume, en iPhone y Android.

El puesto

Line 1: This role targets building the next generation of ML model serving with a focus on reliability, efficiency, and cost-aware scaling.
Line 2: Design and implement real-time inference systems capable of handling high QPS with strict latency SLOs.
Line 3: Operationalize modern inference optimizations—caching, batching, quantization, and memory-efficient techniques—across a diverse hardware fleet.
Line 4: Create platform-wide abstractions to enable broad reuse across workloads and improve developer velocity, observability, and maintainability.
Line 5: Collaborate with ML engineers, infrastructure teams, and OSS communities, contributing back where valuable and influencing the roadmap.
Line 6: Ideal candidates have 8+ years building large-scale serving systems and a track record of mentoring peers and delivering high-quality, scalable software.

Ver la oferta completa

Funciones, perfil, competencias y ventajas — crea tu cuenta gratis.

6 caracteres como mínimo. Cuanto más largo, más seguro.
o

¿Ya tiene una cuenta?

Estas ofertas podrían interesarte

Ninguna oferta realmente parecida por ahora — estas son las más recientes.

Su ubicación

Las ofertas y empresas se filtrarán por este país.

Sugeridos

Todos los países 66