Data Engineer
职位介绍
Architect and build the data infrastructure powering crawling, embedding training, and real-time search.
Design scalable lakehouse architectures (Delta Lake, Iceberg, Hudi) and large distributed pipelines.
Develop streaming systems (Kafka, Flink) and production-grade data processing with Ray, Spark, or ClickHouse.
Prioritize reliability and uptime, delivering systems that rarely page at 3am.
Leverage GPU-accelerated processing (RAPIDS, cuDF) and vector-native storage formats (Lance) where applicable.
Own the data layer for embedding training and indexing workflows at multi-petabyte scales.
Design scalable lakehouse architectures (Delta Lake, Iceberg, Hudi) and large distributed pipelines.
Develop streaming systems (Kafka, Flink) and production-grade data processing with Ray, Spark, or ClickHouse.
Prioritize reliability and uptime, delivering systems that rarely page at 3am.
Leverage GPU-accelerated processing (RAPIDS, cuDF) and vector-native storage formats (Lance) where applicable.
Own the data layer for embedding training and indexing workflows at multi-petabyte scales.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录相似职位
其他可能适合您的职位。