Machine Learning Data Engineer, Replica Pipelines
职位介绍
Design and scale data pipelines for ML datasets used in autonomous systems and simulation environments.
Own data ingestion by building reliable pipelines to normalize and validate customer and synthetic inputs.
Define data standards with schemas, validation checks, and quality metrics for datasets.
Build or fine-tune curation tooling for dataset filtering, versioning, and annotation support.
Enable ML workflows by delivering high-quality data feeds for training and evaluation, in collaboration with ML engineers.
Required: strong Python, data engineering experience, ML-aware engineering, and 3D/vision fundamentals; cloud familiarity and MLOps exposure are valued.
Own data ingestion by building reliable pipelines to normalize and validate customer and synthetic inputs.
Define data standards with schemas, validation checks, and quality metrics for datasets.
Build or fine-tune curation tooling for dataset filtering, versioning, and annotation support.
Enable ML workflows by delivering high-quality data feeds for training and evaluation, in collaboration with ML engineers.
Required: strong Python, data engineering experience, ML-aware engineering, and 3D/vision fundamentals; cloud familiarity and MLOps exposure are valued.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
已有账户? 登录
相似职位
其他可能适合您的职位。
远程办公Partial
城市Vancouver, 加拿大