Vai al contenuto
Whileresume
Classifica Crea il mio CV Assumere Accedi

Staff ML Serving Platform Engineer

Questa offerta fa per te?

Crea il tuo CV e scopri la tua percentuale di corrispondenza con questa posizione — e con tutte le altre.

Crea il mio CV
Segui le tue candidature da mobile L'app gratuita Whileresume, su iPhone e Android.

La posizione

Line 1: This role targets building the next generation of ML model serving with a focus on reliability, efficiency, and cost-aware scaling.
Line 2: Design and implement real-time inference systems capable of handling high QPS with strict latency SLOs.
Line 3: Operationalize modern inference optimizations—caching, batching, quantization, and memory-efficient techniques—across a diverse hardware fleet.
Line 4: Create platform-wide abstractions to enable broad reuse across workloads and improve developer velocity, observability, and maintainability.
Line 5: Collaborate with ML engineers, infrastructure teams, and OSS communities, contributing back where valuable and influencing the roadmap.
Line 6: Ideal candidates have 8+ years building large-scale serving systems and a track record of mentoring peers and delivering high-quality, scalable software.

Vedi l'annuncio completo

Mansioni, profilo, competenze e vantaggi — crea il tuo account gratuito.

Almeno 6 caratteri. Più è lunga, più è sicura.
o

Hai già un account?

Queste offerte potrebbero interessarti

Nessuna offerta davvero simile per ora — ecco le più recenti.

La tua località

Offerte e aziende saranno filtrate su questo paese.

Suggeriti

Tutti i paesi 68