Principal Site Reliability Engineer

GitLab
United States
5 days ago
Apply on job-boards.greenhouse.io
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$223,200.0
Working hours
Regular working hours

Tech stack

Software as a Service Cloud Computing Distributed Systems Python (Programming Language) Reliability Engineering Ruby Backend Gitlab Enterprise Integration Golang

Job description

We’re looking for a Principal Engineer with deep expertise in Site Reliability, Backend, or Platform Engineering to help shape the next phase of GitLab Dedicated, our fully managed single-tenant SaaS offering. This highly influential technical leadership role will set direction for how we scale a growing fleet of isolated, customer-specific environments while maintaining the reliability, security, and compliance our customers depend on.

You’ll lead platform and operating-model transformation across resilience and failover, tenant orchestration, change management, self-service tooling, and platform integrations. You’ll also help align Dedicated with GitLab’s evolution toward more modular and cell-based architectures, strengthen service ownership, and establish scalable patterns that reduce operational complexity. As a Principal Engineer, you’ll influence across teams, guide complex technical decisions, mentor senior engineers, and help raise the technical maturity of the platform as we scale., * Set technical direction for GitLab Dedicated, shaping architecture and platform strategy as we scale a growing fleet of isolated, single-tenant environments.

  • Lead platform transformations across resilience, failover, tenant orchestration, change management, self-service tooling, and platform integrations.
  • Drive scalable, modular architecture that aligns Dedicated with GitLab’s broader Cells strategy while preserving its security, isolation, and compliance requirements.
  • Strengthen service ownership and operational maturity, helping engineering teams build, operate, and improve the production systems they own.
  • Identify and address systemic reliability and scalability risks using production signals, incident patterns, and architectural insight.
  • Establish reusable platform patterns and automation that reduce operational toil and allow Dedicated to scale efficiently.
  • Lead complex technical decisions across teams, balancing reliability, security, cost, maintainability, and customer needs.
  • Advance engineering excellence across the organization through architectural leadership, mentorship, and influence with senior engineers and engineering leaders., * Deep expertise in Site Reliability, Platform, Infrastructure, or Backend Engineering, with experience designing and operating large-scale production systems.

Requirements

  • Hands-on experience with cloud infrastructure, automation, observability, infrastructure as code, and modern production engineering practices.
  • Strong software engineering fundamentals, with experience building production systems or infrastructure tooling in languages such as Go, Ruby, Python, or similar.
  • Strong distributed systems and systems-design expertise, with sound judgment around reliability, failure isolation, scalability, and operational complexity.
  • A track record of technical leadership across multiple teams, setting direction and driving complex initiatives through influence.
  • Experience leading significant platform or infrastructure transformations, including modernization, modularization, or scaling systems through major growth.
  • Experience leading changes that improve how engineering teams own and operate production systems, strengthening reliability, operational readiness, and accountability at scale.
  • Exceptional technical communication and influence, with the ability to build alignment, mentor senior engineers, and guide complex architectural decisions.

The base salary range for this role’s listed level is currently for residents of the United States only. This range is intended to reflect the role’s base salary rate in locations throughout the US. Grade level and salary ranges are determined through interviews and a review of education, experience, knowledge, skills, abilities of the applicant, equity with other team members, alignment with market data, and geographic location. The base salary range does not include any bonuses, equity, or benefits. See more information on our benefits and equity. Sales roles are also eligible for incentive pay targeted at up to 100% of the offered base salary.

About the company

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster.

The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software.

*Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on job-boards.greenhouse.io
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

50 sec

Why developer happiness matters in web frameworks

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

6:14 min

Structuring CI/CD pipelines with integrated security and quality checks

Christoph Ruggenthaler · LIVE

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

5:35 min

Embracing salary transparency in technical job listings

Rudi Bauer Rudi Bauer +2 · Cappuccino with HR

3:30 min

Falling in love with Ruby and creating Basecamp

David Heinemeier Hansson David Heinemeier Hansson +1 · Coffee With Developers

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

Videos

See all

Related articles

See all