Machine Learning Operations Lead
Cette offre est-elle pour vous ?
Créer mon CV Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.
Le poste
Lead the design, delivery, and operation of production ML inference and fine-tuning services across serverless and multi-cluster deployments.
Own availability and performance SLAs, drive incident response, and conduct postmortems to prevent recurrence.
Build and scale testing, deployment, configuration management, and monitoring practices in collaboration with Infra SREs.
Define and enforce configuration best practices for inference engines (vLLM, tvLLM, Pulsar) to prevent runtime issues.
Develop self-serve tooling and internal developer platforms to reduce operational toil for ML engineers and customers.
Lead, mentor, and grow an MLOps team while partnering with infrastructure and ML engineering to improve reliability and cost efficiency.
Own availability and performance SLAs, drive incident response, and conduct postmortems to prevent recurrence.
Build and scale testing, deployment, configuration management, and monitoring practices in collaboration with Infra SREs.
Define and enforce configuration best practices for inference engines (vLLM, tvLLM, Pulsar) to prevent runtime issues.
Develop self-serve tooling and internal developer platforms to reduce operational toil for ML engineers and customers.
Lead, mentor, and grow an MLOps team while partnering with infrastructure and ML engineering to improve reliability and cost efficiency.
Voir l'offre en entier
Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.
ou
Déjà un compte ?
Se connecterOffres similaires
D'autres postes qui pourraient vous convenir.
TélétravailPartial
VilleSan Francisco, États-Unis