ML Engineer - Inference Serving
Is deze vacature iets voor u?
Mijn cv maken Maak uw cv en ontdek uw matchpercentage met deze functie — en met alle andere.
De functie
As an ML Engineer - Inference Serving, you will ship new model architectures by integrating them into our inference engine. Collaborate across research, engineering, and infrastructure to optimize model efficiency and deployment at scale. Build internal tooling to measure, profile, and track the lifetime of inference jobs and workflows. Automate, test, and maintain inference services to ensure maximum uptime and reliability. Design deployment workflows and sophisticated scheduling to optimally leverage GPU resources while meeting SLOs. Develop CI/CD pipelines for processing model checkpoints, platform components, and internal SDKs.
Bekijk de volledige vacature
Taken, profiel, vaardigheden en voordelen — maak gratis een account aan.
Al een account? Inloggen
Thuiswerkenno
StadPalo Alto, US