Machine Learning Data Engineer, Replica Pipelines
仕事内容
As a Machine Learning Data Engineer for Replica Pipelines, you will design and scale data ingestion and processing pipelines for ML workflows.
You will normalize and validate customer and synthetic data to produce reliable, structured feeds for training, evaluation, and production.
Define data standards by creating schemas, validation checks, and quality metrics across Replica datasets.
Build curation tooling for dataset filtering, versioning, and annotation support to enable reproducible data pipelines.
Collaborate with ML engineers to understand data needs and optimize delivery for experimentation and deployment.
Required skills include Python, handling large datasets, 3D vision concepts, and familiarity with cloud and MLOps practices.
You will normalize and validate customer and synthetic data to produce reliable, structured feeds for training, evaluation, and production.
Define data standards by creating schemas, validation checks, and quality metrics across Replica datasets.
Build curation tooling for dataset filtering, versioning, and annotation support to enable reproducible data pipelines.
Collaborate with ML engineers to understand data needs and optimize delivery for experimentation and deployment.
Required skills include Python, handling large datasets, 3D vision concepts, and familiarity with cloud and MLOps practices.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
すでにアカウントをお持ちですか? ログイン
類似の求人
あなたに合いそうな他の職種。
リモートワークPartial
勤務地Vancouver, カナダ