Machine Learning Data Engineer, Replica Pipelines
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Own data ingestion: Build reliable pipelines to normalize and validate customer and synthetic data.
Define data standards: Create schemas, validation checks, and quality metrics for datasets used in ML.
Build curation tooling: Implement tools for dataset filtering, versioning, and annotation support.
Enable ML workflows: Generate high-quality data feeds for training and evaluation across models.
Collaborate closely with ML engineers, applying 3D and computer vision fundamentals to ensure data is ML-ready.
Operate in a modern cloud-based data platform with emphasis on scalability, reproducibility, and dataset governance.
Define data standards: Create schemas, validation checks, and quality metrics for datasets used in ML.
Build curation tooling: Implement tools for dataset filtering, versioning, and annotation support.
Enable ML workflows: Generate high-quality data feeds for training and evaluation across models.
Collaborate closely with ML engineers, applying 3D and computer vision fundamentals to ensure data is ML-ready.
Operate in a modern cloud-based data platform with emphasis on scalability, reproducibility, and dataset governance.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
Already have an account? Log in
Similar openings
Other roles that could suit you.
? Senior Machine Learning Engineer, Pose Estimation ? Senior Machine Learning Engineer, Pose Estimation ? Machine Learning Intern Fall 2026 (Toronto) ? Machine Learning Engineer, Behavior Internship ? Machine Learning Engineer (L3) ? Senior C++ Programmer - Machine Learning Content Creation Technology Group
Remote workPartial
CityVancouver; Karlsruhe, Canada and Germany