Senior Software Engineer, ML Ops & Infrastructure
仕事内容
Design and implement scalable ML infrastructure to train, evaluate, and deploy deep learning models that feed a real-time robotic control stack.
Scale data loading and training across 1000+ GPUs, optimizing throughput and latency.
Build distributed data pipelines for robotics data, enabling efficient model training.
Develop APIs and tooling to let internal and external researchers integrate ML techniques, including open-source model contributions.
Optimize compute resource allocation (GPUs/TPUs) and orchestrate jobs on GKE to reduce cost and improve reliability.
Create model understanding and analysis tools to ensure reliability and traceability across the ML lifecycle.
Scale data loading and training across 1000+ GPUs, optimizing throughput and latency.
Build distributed data pipelines for robotics data, enabling efficient model training.
Develop APIs and tooling to let internal and external researchers integrate ML techniques, including open-source model contributions.
Optimize compute resource allocation (GPUs/TPUs) and orchestrate jobs on GKE to reduce cost and improve reliability.
Create model understanding and analysis tools to ensure reliability and traceability across the ML lifecycle.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログイン類似の求人
あなたに合いそうな他の職種。
? Solutions Architect, Generative AI Deployment ? Senior Consultant - Data Analytics and Supply Chain Transformation ? Dual Study Program in Data Science – in Cooperation with DHBW Stuttgart ? Dual Study Program in Data Science – in cooperation with DHBW Stuttgart ? AI Research Scientist, Vision-guided robotics ? Senior Software Engineer, ML Ops & Infrastructure
リモートワークOn-site
勤務地Munich, ドイツ