Principal Site Reliability Engineer - Austin, Texas

ZOWTA, LLC
Austin, TX, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Compensation
$93,309.0 - $143,208.0
Working hours
Regular working hours

Tech stack

Agile Methodology Artificial Intelligence Amazon Web Services Software as a Service Cloud Computing Cloud Computing Security Cloud Engineering DevOps Linux System Administration Reliability Engineering Software Engineering Data Logging
+9 more
Cloud Platform System System Availability Gitlab Containerization Kubernetes Production Code Build Tools Terraform Dynatrace

Job description

ShipperHQ is looking for a Principal Site Reliability Engineer to lead the evolution of our cloud platform, reliability strategy, and infrastructure architecture. This is a highly technical, hands-on leadership role responsible for designing scalable, resilient systems while establishing engineering best practices that enable our teams to move quickly and confidently. As a Principal SRE, you’ll own the strategic direction of our cloud infrastructure, deployment architecture, observability, and platform reliability. You’ll partner closely with Engineering, Product, Security, and QA to build systems that are secure, automated, highly available, and built to scale. Success in this role comes from balancing strategic thinking with execution and leading through influence, solving complex technical challenges, and continuously improving the developer experience. This role is ideal for someone who enjoys building platforms rather than simply maintaining infrastructure and thrives in a fast-paced, AI-first engineering culture.

  • Own the technical vision and roadmap for ShipperHQ’s cloud infrastructure, reliability, and platform engineering initiatives.
  • Design, build, and maintain highly available, scalable, and secure cloud infrastructure in AWS.
  • Architect and evolve Infrastructure as Code (Terraform) standards across all environments.
  • Design and optimize CI/CD pipelines that enable fast, reliable, and repeatable software delivery.
  • Define and implement reliability standards, SLOs, SLIs, error budgets, and incident management best practices.
  • Lead the design and implementation of observability, monitoring, logging, and alerting across the platform.
  • Build self-service platform capabilities and automation that empower engineering teams and reduce operational overhead.
  • Drive infrastructure modernization initiatives, including containerization, orchestration, and platform scalability.
  • Partner with Security to implement cloud security best practices, compliance controls, and governance.
  • Collaborate with Engineering teams to improve application reliability, performance, and operational excellence.
  • Lead technical decision-making for infrastructure architecture and serve as a trusted advisor across engineering.
  • Mentor engineers and promote best practices in cloud architecture, automation, reliability, and operational excellence.
  • Evaluate and introduce new technologies that improve scalability, reliability, developer productivity, and operational efficiency.
  • Participate in incident response, root cause analysis, and continuous improvement efforts for production systems.

Requirements

  • 10+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, Cloud Infrastructure, or Software Engineering.
  • Proven experience designing and operating large-scale, highly available cloud infrastructure in AWS.
  • Strong software engineering background with the ability to write production-quality code and automation.
  • Expert-level experience with Infrastructure as Code, preferably Terraform.
  • Deep experience designing and maintaining modern CI/CD pipelines using GitLab or similar platforms.
  • Strong knowledge of Kubernetes, containerized workloads, and cloud-native architectures.
  • Extensive experience with observability platforms, distributed tracing, logging, monitoring, and incident response.
  • Experience defining and implementing SLOs, SLIs, and reliability engineering best practices.
  • Strong understanding of networking, security, Linux systems administration, and cloud architecture.
  • Experience supporting high-traffic SaaS applications and mission-critical production environments.
  • Excellent problem-solving skills with the ability to simplify complex technical challenges.
  • Demonstrated ability to influence technical direction without direct authority while mentoring engineers across multiple teams.
  • Experience working in Agile development environments and partnering closely with cross-functional engineering teams.

Benefits & conditions

This is a highly fast-paced environment where no two days will look alike. For the right candidate, with the right attitude, there are fantastic opportunities for career progression. We are an agile, fast-moving team that likes to roll up our sleeves and solve some of the biggest issues in shipping. You will learn more at ShipperHQ in a year than you would in 3 years at other companies, thanks to our collaborative learning culture that fosters continuous growth and innovation. Benefits and Perks:

  • Collaborate with a motivated team, directly tying your results to organizational success
  • 22 days of PTO plus public holidays
  • 401k Match
  • Medical, Dental, and Vision Insurance
  • Maternity and Paternity Leave
  • This is a hybrid, full-time position working out of our Austin, TX office in the Arboretum Area
  • Compensation is based on experience

At ShipperHQ, we’re proud to be a team that’s as diverse as the merchants we serve. As a member of the e-commerce community, we take responsibility to empower shops large and small to grow and thrive through the power of technology to heart. With honesty, responsiveness, and innovation at the center of all we do, we remain committed to hiring the right people for the job, regardless of race, background, religion, or eccentricity. Powered by JazzHR

About the company

ShipperHQ is a trusted leader in the e-commerce shipping space, with over 15 years of experience helping merchants deliver better checkout experiences. Founded in 2009, we power shipping logic and checkout optimization for thousands of brands, from DTC disruptors to enterprise retailers, in 150+ countries. Based in Austin with a global team, we’re a fast-moving, product-led company shaping the future of e-commerce logistics., Realtor.com

  • Austin, TX Recognized as the No. 1 site trusted by real estate professionals, Realtor.com® has been at the forefront of online real estate for over 25 years, connecting buyers, sellers, and r…

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · WWC 2021

1:01 min

Connecting frontend application performance to user retention and revenue

Dani Coll Dani Coll · WWC 2025

6:14 min

Structuring CI/CD pipelines with integrated security and quality checks

Christoph Ruggenthaler · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

12:08 min

Comparing Keptn orchestration capabilities against alternative software operators

Thomas Schütz · LIVE

Videos

See all

Related articles

See all