Sr. Cloud Operations Reliability Engineer (SRE)

nextgen · Remote

Posted
30 days ago
Last confirmed live
Today

What this role involves

The Senior Cloud Operations Reliability Engineer is responsible for driving operational excellence and strengthening the reliability posture of cloud-based services. This role involves establishing observability practices, leading incident response, and implementing automation to reduce toil. The position requires 10+ years of experience in cloud operations or SRE, with expertise in GCP or AWS and Infrastructure as Code.

Skills this posting asks for

  • google cloud platform
  • aws
  • terraform
  • deployment manager
  • cloudformation
  • infrastructure as code
  • monitoring
  • observability
  • alerting
  • incident response
  • root cause analysis
  • slo
  • sli
  • automation
  • runbooks
  • logging
  • capacity planning
  • disaster recovery
  • business continuity
  • cloud governance
  • compliance
  • security
  • access control
  • tagging

Requirements

  • 10 years of experience
  • Level: senior
  • Remote policy: remote

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at nextgen

All 3 roles at nextgen