Data Engineer
Is deze vacature iets voor u?
Mijn cv maken Maak uw cv en ontdek uw matchpercentage met deze functie — en met alle andere.
De functie
Design, build, and operate batch data pipelines using Python, Spark (EMR), and AWS Glue to load data from S3, RDS, and external sources into Hive/Athena tables.
Model datasets in the S3/Hive data lake to support analytics, API use cases, Elasticsearch indexes, and ML models.
Implement and run Airflow (MWAA) workflows with dependency management, scheduling, retries, and alerting via Slack.
Build robust data quality checks and validation, with monitoring and quick issue surface.
Optimize jobs for cost and performance through partitioning, file formats, and efficient resource usage.
Collaborate with data scientists, ML engineers, and application engineers to design schemas and pipelines that serve multiple use cases and contribute to tooling and best practices.
Model datasets in the S3/Hive data lake to support analytics, API use cases, Elasticsearch indexes, and ML models.
Implement and run Airflow (MWAA) workflows with dependency management, scheduling, retries, and alerting via Slack.
Build robust data quality checks and validation, with monitoring and quick issue surface.
Optimize jobs for cost and performance through partitioning, file formats, and efficient resource usage.
Collaborate with data scientists, ML engineers, and application engineers to design schemas and pipelines that serve multiple use cases and contribute to tooling and best practices.
Bekijk de volledige vacature
Taken, profiel, vaardigheden en voordelen — maak gratis een account aan.
of
Al een account?
InloggenVergelijkbare vacatures
Andere functies die kunnen passen.