Data Engineer
Is deze vacature iets voor u?
Mijn cv maken Maak uw cv en ontdek uw matchpercentage met deze functie — en met alle andere.
De functie
Architect and build the data infrastructure powering crawling billions of pages, embedding model training, and real-time search.
Design lakehouse architectures using Delta Lake, Iceberg, or Hudi and decide when to use each.
Build and operate large-scale distributed data processing pipelines and streaming systems (Kafka, Flink).
Work with Ray, Spark, or ClickHouse in production to enable fast analytics and indexing pipelines.
Maintain a relentless focus on reliability to ensure systems rarely require on-call support.
Scale data infrastructure to hundreds of petabytes and own the end-to-end data layer for embedding training and real-time indexing.
Design lakehouse architectures using Delta Lake, Iceberg, or Hudi and decide when to use each.
Build and operate large-scale distributed data processing pipelines and streaming systems (Kafka, Flink).
Work with Ray, Spark, or ClickHouse in production to enable fast analytics and indexing pipelines.
Maintain a relentless focus on reliability to ensure systems rarely require on-call support.
Scale data infrastructure to hundreds of petabytes and own the end-to-end data layer for embedding training and real-time indexing.
Bekijk de volledige vacature
Taken, profiel, vaardigheden en voordelen — maak gratis een account aan.
of
Al een account?
InloggenVergelijkbare vacatures
Andere functies die kunnen passen.