Machine Learning Data Engineer, Replica Pipelines
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Design and scale data pipelines for ML datasets used in autonomous systems and simulation environments.
Own data ingestion by building reliable pipelines to normalize and validate customer and synthetic inputs.
Define data standards with schemas, validation checks, and quality metrics for datasets.
Build or fine-tune curation tooling for dataset filtering, versioning, and annotation support.
Enable ML workflows by delivering high-quality data feeds for training and evaluation, in collaboration with ML engineers.
Required: strong Python, data engineering experience, ML-aware engineering, and 3D/vision fundamentals; cloud familiarity and MLOps exposure are valued.
Own data ingestion by building reliable pipelines to normalize and validate customer and synthetic inputs.
Define data standards with schemas, validation checks, and quality metrics for datasets.
Build or fine-tune curation tooling for dataset filtering, versioning, and annotation support.
Enable ML workflows by delivering high-quality data feeds for training and evaluation, in collaboration with ML engineers.
Required: strong Python, data engineering experience, ML-aware engineering, and 3D/vision fundamentals; cloud familiarity and MLOps exposure are valued.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
Already have an account? Log in
Similar openings
Other roles that could suit you.
Remote workPartial
CityVancouver, Canada