Senior Data Engineer - Product
职位介绍
Re-architect and scale the DSF-based big data processing components powering production risk analytics.
Analyze workload patterns across Spark jobs, notebooks, and DS API usage to drive performance, reliability, and cost improvements.
Ensure stability of Spark jobs running on EMR or Kubernetes and operate Hadoop ecosystem components (HDFS, YARN) and Spark runtimes.
Maintain and evolve ingestion pipelines between Runtime and DSF (Firehose, Glue → S3) and improve Spark runtimes.
Enhance developer and power-user experience across JupyterLabs and the DS API.
Own services throughout their lifecycle following DevOps practices, collaborating with data scientists and platform teams to deliver a reliable, scalable data platform.
Analyze workload patterns across Spark jobs, notebooks, and DS API usage to drive performance, reliability, and cost improvements.
Ensure stability of Spark jobs running on EMR or Kubernetes and operate Hadoop ecosystem components (HDFS, YARN) and Spark runtimes.
Maintain and evolve ingestion pipelines between Runtime and DSF (Firehose, Glue → S3) and improve Spark runtimes.
Enhance developer and power-user experience across JupyterLabs and the DS API.
Own services throughout their lifecycle following DevOps practices, collaborating with data scientists and platform teams to deliver a reliable, scalable data platform.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录相似职位
其他可能适合您的职位。
远程办公Full remote
城市Remote / Global, Worldwide