Data Engineer
职位介绍
Architect and build the data infrastructure powering crawling billions of pages, embedding model training, and real-time search.
Design lakehouse architectures using Delta Lake, Iceberg, or Hudi and decide when to use each.
Build and operate large-scale distributed data processing pipelines and streaming systems (Kafka, Flink).
Work with Ray, Spark, or ClickHouse in production to enable fast analytics and indexing pipelines.
Maintain a relentless focus on reliability to ensure systems rarely require on-call support.
Scale data infrastructure to hundreds of petabytes and own the end-to-end data layer for embedding training and real-time indexing.
Design lakehouse architectures using Delta Lake, Iceberg, or Hudi and decide when to use each.
Build and operate large-scale distributed data processing pipelines and streaming systems (Kafka, Flink).
Work with Ray, Spark, or ClickHouse in production to enable fast analytics and indexing pipelines.
Maintain a relentless focus on reliability to ensure systems rarely require on-call support.
Scale data infrastructure to hundreds of petabytes and own the end-to-end data layer for embedding training and real-time indexing.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录相似职位
其他可能适合您的职位。