Skip to content

Machine Learning Data Engineer, Replica Pipelines

Is this job for you?

Build your CV and see how well you match this role — and every other one.

Build my CV

The role

Own data ingestion and build reliable pipelines to normalize and validate customer and synthetic data for Replica and ML model development.
Define data standards by creating schemas, validation checks, and quality metrics to ensure clean, consistent datasets.
Develop curation tooling for dataset filtering, versioning, and annotation support to enable efficient data governance.
Enable ML workflows by generating high-quality data feeds for training and evaluation across models.
Collaborate closely with ML engineers to translate data needs into scalable, production-ready data infrastructure.
Leverage Python, cloud storage, and 3D/vision fundamentals to manage large-scale datasets used in autonomous systems.

See the full job post

Responsibilities, requirements, skills and benefits — create your free account.

At least 8 characters, one uppercase letter and one digit.

Already have an account? Log in

Similar openings

Other roles that could suit you.

See all →

Your location

Jobs and companies will be filtered on this country.

Suggested

All countries 65