Machine Learning Data Engineer, Replica Pipelines
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
Track your applications on mobile The free Whileresume app, on iPhone and Android.
The role
Own data ingestion: Build reliable pipelines to normalize and validate customer and synthetic data.
Define data standards: Create schemas, validation checks, and quality metrics for datasets used in ML.
Build curation tooling: Implement tools for dataset filtering, versioning, and annotation support.
Enable ML workflows: Generate high-quality data feeds for training and evaluation across models.
Collaborate closely with ML engineers, applying 3D and computer vision fundamentals to ensure data is ML-ready.
Operate in a modern cloud-based data platform with emphasis on scalability, reproducibility, and dataset governance.
Define data standards: Create schemas, validation checks, and quality metrics for datasets used in ML.
Build curation tooling: Implement tools for dataset filtering, versioning, and annotation support.
Enable ML workflows: Generate high-quality data feeds for training and evaluation across models.
Collaborate closely with ML engineers, applying 3D and computer vision fundamentals to ensure data is ML-ready.
Operate in a modern cloud-based data platform with emphasis on scalability, reproducibility, and dataset governance.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inYou might also like these jobs
No closely matching jobs yet — here are the most recent ones.