Data Engineer
Is deze vacature iets voor u?
Mijn cv maken Maak uw cv en ontdek uw matchpercentage met deze functie — en met alle andere.
De functie
Architect and build the data infrastructure powering crawling, embedding training, and real-time search.
Design scalable lakehouse architectures (Delta Lake, Iceberg, Hudi) and large distributed pipelines.
Develop streaming systems (Kafka, Flink) and production-grade data processing with Ray, Spark, or ClickHouse.
Prioritize reliability and uptime, delivering systems that rarely page at 3am.
Leverage GPU-accelerated processing (RAPIDS, cuDF) and vector-native storage formats (Lance) where applicable.
Own the data layer for embedding training and indexing workflows at multi-petabyte scales.
Design scalable lakehouse architectures (Delta Lake, Iceberg, Hudi) and large distributed pipelines.
Develop streaming systems (Kafka, Flink) and production-grade data processing with Ray, Spark, or ClickHouse.
Prioritize reliability and uptime, delivering systems that rarely page at 3am.
Leverage GPU-accelerated processing (RAPIDS, cuDF) and vector-native storage formats (Lance) where applicable.
Own the data layer for embedding training and indexing workflows at multi-petabyte scales.
Bekijk de volledige vacature
Taken, profiel, vaardigheden en voordelen — maak gratis een account aan.
of
Al een account?
InloggenVergelijkbare vacatures
Andere functies die kunnen passen.