Sr. Cloud Operations Reliability Engineer (SRE)
nextgen · Remote
- Posted
- 30 days ago
- Last confirmed live
- Today
What this role involves
The Senior Cloud Operations Reliability Engineer is responsible for driving operational excellence and strengthening the reliability posture of cloud-based services. This role involves establishing observability practices, leading incident response, and implementing automation to reduce toil. The position requires 10+ years of experience in cloud operations or SRE, with expertise in GCP or AWS and Infrastructure as Code.
Skills this posting asks for
- google cloud platform
- aws
- terraform
- deployment manager
- cloudformation
- infrastructure as code
- monitoring
- observability
- alerting
- incident response
- root cause analysis
- slo
- sli
- automation
- runbooks
- logging
- capacity planning
- disaster recovery
- business continuity
- cloud governance
- compliance
- security
- access control
- tagging
Requirements
- 10 years of experience
- Level: senior
- Remote policy: remote
Apply without filling the form
Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.