Senior Software Engineer (MLOps) – Annotation & Evaluation
Cette offre est-elle pour vous ?
Créer mon CV Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.
Suivez vos candidatures sur mobile L'application Whileresume, gratuite, sur iPhone et Android.
Le poste
Design and build systems for automated evaluation of AI models, including LLMs and agents, using production-like telemetry and realistic scenarios.
Lead the development of benchmark suites, evaluation pipelines, and model comparison tools with integrated trust & safety metrics.
Build and maintain integrations with labeling systems (e.g., Label Studio) and coordinate with external/internal annotation workflows.
Collaborate with Applied AI and Bits AI teams to enable fast iteration, reproducible experiments, and interpretable evaluations.
Develop data pipelines that feed metrics, results, and alerts into our observability stack to monitor model behavior at scale.
Promote safe deployment practices via bias checks, hallucination detection, and human-in-the-loop review.
Lead the development of benchmark suites, evaluation pipelines, and model comparison tools with integrated trust & safety metrics.
Build and maintain integrations with labeling systems (e.g., Label Studio) and coordinate with external/internal annotation workflows.
Collaborate with Applied AI and Bits AI teams to enable fast iteration, reproducible experiments, and interpretable evaluations.
Develop data pipelines that feed metrics, results, and alerts into our observability stack to monitor model behavior at scale.
Promote safe deployment practices via bias checks, hallucination detection, and human-in-the-loop review.
Voir l'offre en entier
Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.
ou
Déjà un compte ?
Se connecterCes offres pourraient vous intéresser
Aucune offre vraiment proche pour l'instant — voici les plus récentes.