Naar de inhoud
Whileresume
Klassement Mijn cv maken Werven Inloggen

Staff ML Serving Platform Engineer

Is deze vacature iets voor u?

Maak uw cv en ontdek uw matchpercentage met deze functie — en met alle andere.

Mijn cv maken
Volg uw sollicitaties op mobiel De gratis Whileresume-app, op iPhone en Android.

De functie

Line 1: This role targets building the next generation of ML model serving with a focus on reliability, efficiency, and cost-aware scaling.
Line 2: Design and implement real-time inference systems capable of handling high QPS with strict latency SLOs.
Line 3: Operationalize modern inference optimizations—caching, batching, quantization, and memory-efficient techniques—across a diverse hardware fleet.
Line 4: Create platform-wide abstractions to enable broad reuse across workloads and improve developer velocity, observability, and maintainability.
Line 5: Collaborate with ML engineers, infrastructure teams, and OSS communities, contributing back where valuable and influencing the roadmap.
Line 6: Ideal candidates have 8+ years building large-scale serving systems and a track record of mentoring peers and delivering high-quality, scalable software.

Bekijk de volledige vacature

Taken, profiel, vaardigheden en voordelen — maak gratis een account aan.

Minimaal 6 tekens. Hoe langer, hoe veiliger.
of

Al een account?

Deze vacatures zijn misschien iets voor u

Nog geen echt vergelijkbare vacatures — hier zijn de nieuwste.

Uw locatie

Vacatures en bedrijven worden op dit land gefilterd.

Voorgesteld

Alle landen 68