Senior Machine Learning Engineer - Evaluation
Ist diese Stelle etwas für Sie?
Lebenslauf erstellen Erstellen Sie Ihren Lebenslauf und entdecken Sie Ihre Übereinstimmung mit dieser Stelle — und mit allen anderen.
Die Stelle
Lead the design and implementation of offline evaluation pipelines for embodied AI models, building scalable and interpretable benchmarks that cover action, perception, and language-grounded reasoning. Design metrics across vision, language, and driving tasks to robustly evaluate model behavior and inform deployment readiness. Drive human annotation workflows, including task design, QA, and coordination with internal teams and external partners, to ensure high-quality ground-truth data. Collaborate with science, datasets, and infrastructure teams to align evaluation with product goals and system safety. Analyze offline metrics and correlate them with online performance to guide model selection and risk assessment. Demonstrate strong Python software engineering and data processing skills to deliver scalable, reproducible evaluation tooling.
Die vollständige Anzeige sehen
Aufgaben, Anforderungen, Kompetenzen und Vorteile — mit Ihrem kostenlosen Konto.
oder
Bereits ein Konto?
AnmeldenÄhnliche Stellen
Weitere Positionen, die passen könnten.