Data Engineer
Cette offre est-elle pour vous ?
Créer mon CV Créez votre CV et découvrez votre pourcentage de correspondance avec ce poste — et avec tous les autres.
Le poste
Architect and build the data infrastructure powering crawling, embedding training, and real-time search.
Design scalable lakehouse architectures (Delta Lake, Iceberg, Hudi) and large distributed pipelines.
Develop streaming systems (Kafka, Flink) and production-grade data processing with Ray, Spark, or ClickHouse.
Prioritize reliability and uptime, delivering systems that rarely page at 3am.
Leverage GPU-accelerated processing (RAPIDS, cuDF) and vector-native storage formats (Lance) where applicable.
Own the data layer for embedding training and indexing workflows at multi-petabyte scales.
Design scalable lakehouse architectures (Delta Lake, Iceberg, Hudi) and large distributed pipelines.
Develop streaming systems (Kafka, Flink) and production-grade data processing with Ray, Spark, or ClickHouse.
Prioritize reliability and uptime, delivering systems that rarely page at 3am.
Leverage GPU-accelerated processing (RAPIDS, cuDF) and vector-native storage formats (Lance) where applicable.
Own the data layer for embedding training and indexing workflows at multi-petabyte scales.
Voir l'offre en entier
Missions, profil recherché, compétences et avantages — en créant votre compte gratuitement.
ou
Déjà un compte ?
Se connecterOffres similaires
D'autres postes qui pourraient vous convenir.