Skip to content

Data Engineer

Is this job for you?

Build your CV and see how well you match this role — and every other one.

Build my CV

The role

Design, build, and operate batch data pipelines using Python, Spark (EMR), and AWS Glue to load data from S3, RDS, and external sources into Hive/Athena tables.
Model datasets in the S3/Hive data lake to support analytics, API use cases, Elasticsearch indexes, and ML models.
Implement and run Airflow (MWAA) workflows with dependency management, scheduling, retries, and alerting via Slack.
Build robust data quality checks and validation, with monitoring and quick issue surface.
Optimize jobs for cost and performance through partitioning, file formats, and efficient resource usage.
Collaborate with data scientists, ML engineers, and application engineers to design schemas and pipelines that serve multiple use cases and contribute to tooling and best practices.

See the full job post

Responsibilities, requirements, skills and benefits — create your free account.

At least 6 characters. The longer, the safer.
or

Already have an account?

Similar openings

Other roles that could suit you.

See all →

Your location

Jobs and companies will be filtered on this country.

Suggested

All countries 65