Cloud Site Reliability Engineer
职位介绍
Own and maintain the uptime and reliability of cloud-based services through proactive monitoring, automation, and incident response.
Automate 50% of site systems to self-manage and self-heal for faster recovery and reduced toil.
Own deployment, incident handling, on-call duty, and manual interventions to ensure smooth production operations.
Collaborate with Cloud DevOps to implement robust monitoring, alerting, and governance across public clouds.
Apply scripting and configuration management to drive scalable, repeatable deployments and changes.
Communicate clearly, document changes, and troubleshoot complex issues across Windows and Linux environments.
Automate 50% of site systems to self-manage and self-heal for faster recovery and reduced toil.
Own deployment, incident handling, on-call duty, and manual interventions to ensure smooth production operations.
Collaborate with Cloud DevOps to implement robust monitoring, alerting, and governance across public clouds.
Apply scripting and configuration management to drive scalable, repeatable deployments and changes.
Communicate clearly, document changes, and troubleshoot complex issues across Windows and Linux environments.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
已有账户? 登录
城市Santa Clara, United States