Principal Site Reliability Engineer - Austin, Texas

ShipperHQ · Austin, TX

The role

ShipperHQ seeks a Principal Site Reliability Engineer to own the technical vision for its cloud infrastructure, reliability, and platform engineering. This hands-on leadership role involves designing scalable, resilient systems, establishing engineering best practices, and mentoring engineers. The position is based in Austin, Texas, with a hybrid work model.

Location
Austin, TX
Work mode
Hybrid
Employment
Full time
Level
Principal
Experience
10+ yrs

What we know that the posting doesn’t say

  • Seen 2 days agostill listed on the employer’s careers page
  • Posted 77 days agothe first time we saw it

About ShipperHQ

ShipperHQ is a trusted leader in e-commerce shipping, providing shipping logic and checkout optimization for thousands of brands globally.

What you would do

  • Own the technical vision and roadmap for cloud infrastructure and reliability.
  • Design, build, and maintain highly available, scalable, secure AWS cloud infrastructure.
  • Architect and evolve Infrastructure as Code (Terraform) standards across environments.
  • Design and optimize CI/CD pipelines for fast, reliable software delivery.
  • Define and implement reliability standards, SLOs, SLIs, and incident management practices.
  • Lead design and implementation of observability, monitoring, logging, and alerting.

Must have

  • 10+ years in SRE, Platform Engineering, DevOps, Cloud Infrastructure, or Software Engineering.
  • Proven experience designing large-scale, highly available AWS cloud infrastructure.
  • Strong software engineering background with production-quality code and automation.
  • Expert-level experience with Infrastructure as Code, preferably Terraform.
  • Deep experience designing modern CI/CD pipelines using GitLab or similar.
  • Strong knowledge of Kubernetes, containerized workloads, and cloud-native architectures.
  • Extensive experience with observability, distributed tracing, logging, monitoring, and incident response.
  • Experience defining and implementing SLOs, SLIs, and reliability engineering best practices.
  • Strong understanding of networking, security, Linux systems administration, and cloud architecture.
  • Experience supporting high-traffic SaaS applications and mission-critical production environments.
  • Excellent problem-solving skills and ability to simplify complex technical challenges.
  • Demonstrated ability to influence technical direction and mentor engineers across teams.

What you get

  • 22 days of PTO plus public holidays
  • 401k match
  • Medical, dental, and vision insurance
  • Maternity and paternity leave

Experience and education

  • 10+ years of relevant experience

Key skills

  • aws
  • terraform
  • gitlab
  • kubernetes
  • linux
  • ci/cd
  • observability
  • monitoring
  • logging
  • incident response
  • slo
  • sli
  • error budgets
  • containerization
  • cloud-native architecture
  • networking
  • security
  • agile

Principal Site Reliability Engineer About ShipperHQ: ShipperHQ is a trusted leader in the e-commerce shipping space, with over 15 years of experience helping merchants deliver better checkout experiences. Founded in 2009, we power shipping logic and checkout optimization for thousands of brands, fr…

Extracted from the employer’s posting. Read it in full on ShipperHQ’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at ShipperHQ

All 4 roles at ShipperHQ