Data Engineer
仕事内容
Design, build, and operate batch data pipelines using Python, Spark (EMR), and AWS Glue to load data from S3, RDS, and external sources into Hive/Athena tables.
Model datasets in the S3/Hive data lake to support analytics, API use cases, Elasticsearch indexes, and ML models.
Implement and run Airflow (MWAA) workflows with dependency management, scheduling, retries, and alerting via Slack.
Build robust data quality checks and validation, with monitoring and quick issue surface.
Optimize jobs for cost and performance through partitioning, file formats, and efficient resource usage.
Collaborate with data scientists, ML engineers, and application engineers to design schemas and pipelines that serve multiple use cases and contribute to tooling and best practices.
Model datasets in the S3/Hive data lake to support analytics, API use cases, Elasticsearch indexes, and ML models.
Implement and run Airflow (MWAA) workflows with dependency management, scheduling, retries, and alerting via Slack.
Build robust data quality checks and validation, with monitoring and quick issue surface.
Optimize jobs for cost and performance through partitioning, file formats, and efficient resource usage.
Collaborate with data scientists, ML engineers, and application engineers to design schemas and pipelines that serve multiple use cases and contribute to tooling and best practices.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログイン類似の求人
あなたに合いそうな他の職種。