Aller au contenu

Senior Machine Learning Engineer - Evaluation

Cette offre est-elle pour vous ?

Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.

Créer mon CV

Le poste

Lead the design and implementation of offline evaluation pipelines for embodied AI models, building scalable and interpretable benchmarks that cover action, perception, and language-grounded reasoning. Design metrics across vision, language, and driving tasks to robustly evaluate model behavior and inform deployment readiness. Drive human annotation workflows, including task design, QA, and coordination with internal teams and external partners, to ensure high-quality ground-truth data. Collaborate with science, datasets, and infrastructure teams to align evaluation with product goals and system safety. Analyze offline metrics and correlate them with online performance to guide model selection and risk assessment. Demonstrate strong Python software engineering and data processing skills to deliver scalable, reproducible evaluation tooling.

Voir l'offre en entier

Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.

6 caractères minimum. Plus il est long, plus il est sûr.
ou

Déjà un compte ?

Offres similaires

D'autres postes qui pourraient vous convenir.

Voir tout →

Votre lieu

Les offres et entreprises seront filtrées sur ce pays.

Suggérés

Tous les pays 65