Senior Software Engineer (MLOps) – Serving
Cette offre est-elle pour vous ?
Créer mon CV Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.
Le poste
Architect and build scalable ML/LLM model-serving systems across data centers with strong SLAs, observability, and reliability.
Design and optimize Ray-based inference infrastructure to handle both low- and high-throughput workloads.
Enable applied scientists to deploy and test models via self-service tools, CI/CD pipelines, and rollback mechanisms.
Implement A/B testing and shadow deployment capabilities to evaluate new model versions in production.
Collaborate with platform teams to improve GPU provisioning, traffic routing, and runtime performance.
Instrument inference workflows with telemetry (latency, token counts, errors) to drive performance and safety analysis.
Design and optimize Ray-based inference infrastructure to handle both low- and high-throughput workloads.
Enable applied scientists to deploy and test models via self-service tools, CI/CD pipelines, and rollback mechanisms.
Implement A/B testing and shadow deployment capabilities to evaluate new model versions in production.
Collaborate with platform teams to improve GPU provisioning, traffic routing, and runtime performance.
Instrument inference workflows with telemetry (latency, token counts, errors) to drive performance and safety analysis.
Voir l'offre en entier
Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.
Déjà un compte ? Se connecter
Offres similaires
D'autres postes qui pourraient vous convenir.
? Software Engineer, Gen AI & Web Apps ? Senior Software Engineer - Big Data/GenAI ? Director, Software Engineering — Marketing Intelligence (AI & Agents) ? Principal Software Engineer, ML Flywheel Technical Lead ? Software Engineer (MLOps) - H/F/NB ? Software Engineer, Trustworthy AI - Enterprise AI
Télétravailpartial
VilleFlexible/Remote (France-based), France