> Markdown version of [/jobs/ext/3119625-lead-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3119625-lead-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Site Reliability Engineer - **Company:** CareerCircle - **Location:** Berkeley, MO, United States - **Experience:** Expert - **Salary:** $198,050.0 - $267,950.0 - **Contract:** Permanent contract - **Skills:** Microsoft Access, Java (Programming Language), Agile Methodology, Amazon Web Services, Confluence, JIRA, Microsoft Azure, Backup Devices, Bash Shell, C Sharp (Programming Language), C++ (Programming Language), Cloud Computing, Cloud Engineering, Configuration Management, Code Review, CompTIA Security+, Cyber Security, Databases, Continuous Integration, Database Storage Structures, Decision Support Systems, DevOps, Programming Tools, Disaster Recovery, Information Technology Operations, Python (Programming Language), Key Management, PostgreSQL, Performance Tuning, Windows PowerShell, Reliability Engineering, Ansible, SAP (Applications), Software Deployment, Software Engineering, SonarQube, SQL Databases, System Software, Virtualization Technology, Software Vulnerability Management, Technical Debt, Infrastructure as Code (IaC), Gitlab, Git, SC Clearance, Gitlab-ci, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Atlassian Tools, Patch Management, Enterprise Integration, Hardware Infrastructure, Software Version Control, Devsecops, Atlassian Bamboo, Docker, Jenkins, Artifactory, Programming Languages - **Published:** September 28, 2026 - **Apply:** https://www.careercircle.com/jobs/all/all/usa/mo/berkeley/c90aaf01-67f9-46fd-89bd-c2e0d7b06907 ## About the Role Gitlab Ansible Tooling Physics Jenkins Planning Auditing Topology Dashboard Chemistry Hardening SonarQube Operations Leadership Management Automation Kubernetes PostgreSQL Mathematics Scalability Reliability Artifactory End Systems Data Science GitLab CI/CD Data Privacy Communication Investigation Observability Cyber Security Export Control Virtualization Technical Debt Version Control Microsoft Azure Access Controls Computer Science SAP Applications Secret Clearance Health Insurance Security Controls Docker (Software) CompTIA Security+ Disaster Recovery Influencing Skills Performance Tuning Security Clearance Technical Strategy Root Cause Analysis Amazon Web Services Technical Authority Software Engineering Technical Leadership Atlassian Confluence Collective Bargaining Operational Excellence Enterprise Architecture Configuration Management Manufacturing Engineering Infrastructure Automation Backup and Recovery System Verbal Communication Skills, * Bachelors Degree * This position requires the ability to obtain a US Security Clearance for which the US Government requires US Citizenship * This position requires ability to obtain access to Special Access Programs (SAP) * 8+ years of experience with CI/CD tools such as Jenkins or Bamboo * 8+ years of experience enterprise architecture experience, including but not limited to cloud architecture, security, data privacy, integration, and deployment * 5+ years of experience in root cause analysis and corrective action * 5+ years of technical leadership and team leadership, * Bachelor of Science degree from an accredited course of study in engineering, engineering technology (includes manufacturing engineering technology), chemistry, physics, mathematics, data science, or computer science and 14+ years of related work experience or Bachelor's Degree and 18+ years of directly related work experience or 22+ years of related, relevant experience * Experience effectively communicating technical strategy, risk, tradeoffs, and recommendations to senior technical and program leadership * Vast experience administering or architecting GitLab, GitLab CI/CD, GitLab runners, or comparable enterprise source control and CI/CD platforms * Deep experience administering or architecting Jira, Confluence, or other Atlassian products in an enterprise environment * Deep experience with PostgreSQL architecture and operations, including backup and recovery, replication, performance tuning, storage planning, maintenance, and high-availability patterns * Experience with AWS, Microsoft Azure, Infrastructure as Code, Ansible, configuration management, containers, Docker, Kubernetes, virtualization, artifact management, secrets management, and secure software delivery practices * Experience administering or architecting Artifactory, SonarQube, Jenkins, or similar software delivery tools * Experience designing observability platforms, alerting strategies, SLO frameworks, service health dashboards, and operational reporting * Experience supporting Air Dominance, classified, air-gapped, or highly regulated engineering environments * Experience developing disaster recovery strategy, continuity of operations plans, recovery time objectives, recovery point objectives, and restore validation programs * Experience guiding cybersecurity hardening, vulnerability remediation, audit readiness, privileged access controls, and compliance-driven operations * Ability to obtain Security+ certification * Demonstrated ability to lead through influence across engineering teams, customers, suppliers, cybersecurity, infrastructure, and program stakeholders * Strong written and verbal communication skills with the ability to produce architecture documentation, executive briefings, technical roadmaps, and decision records, This position must meet U.S. export control compliance requirements. To meet U.S. export control compliance requirements, a "U.S. Person" as defined by 22 C.F.R. §120.62 is required. "U.S. Person" includes U.S. Citizen, U.S. National, lawful permanent resident, refugee, or asylee., Bachelor's Degree or Equivalent Required, This position requires an active U.S. Secret Security Clearance (U.S. Citizenship Required). (A U.S. Security Clearance that has been active in the past 24 months is considered active), Ansible Tooling Physics Jenkins Planning Auditing Topology Dashboard Chemistry Hardening SonarQube Operations Leadership Management Automation Kubernetes PostgreSQL Mathematics Scalability Reliability Artifactory End Systems Data Science GitLab CI/CD Data Privacy Communication Investigation Observability Cyber Security Export Control Virtualization Technical Debt Version Control Microsoft Azure Access Controls Computer Science SAP Applications Secret Clearance Health Insurance Security Controls Docker (Software) CompTIA Security+ Disaster Recovery Influencing Skills Performance Tuning Security Clearance Technical Strategy Root Cause Analysis Amazon Web Services Technical Authority Software Engineering Technical Leadership Atlassian Confluence Collective Bargaining Operational Excellence Enterprise Architecture Configuration Management Manufacturing Engineering Infrastructure Automation Backup and Recovery System Verbal Communication Skills ## Description The Boeing Company is looking for a Lead Site Reliability Engineer to join the Air Dominance Site Reliability Engineering team located in Berkeley, MO. We are seeking a highly talented, motivated, and creative technical leader responsible for the reliability strategy, architecture, operational maturity, and long-term technical direction of mission-critical developer platforms used by Air Dominance engineering teams., This role will provide technical leadership across GitLab, GitLab CI/CD runners, Jira, Confluence, PostgreSQL, related software delivery tools such as Artifactory and SonarQube, and the supporting infrastructure, automation, monitoring, backup, recovery, and security controls required to operate these services. The selected candidate will do the following; define standards, guide architecture decisions, mentor engineers, lead complex technical investigations, and partner with program leadership and stakeholders. They will ensure developer tooling remains secure, reliable, scalable, and are able to support., * Define and lead the Site Reliability Engineering technical strategy for GitLab, CI/CD runners, Jira, Confluence, PostgreSQL, Artifactory, SonarQube, and related developer tooling infrastructure * Establish platform reliability architecture, operational standards, SLIs, SLOs, SLAs, KPIs, error budgets, observability patterns, capacity models, backup strategies, and disaster recovery approaches * Serve as the senior technical authority for complex reliability, performance, scalability, integration, database, automation, and security-related platform decisions * Lead architecture and design reviews for developer tooling infrastructure, CI/CD runner topology, PostgreSQL operations, cloud-based and on-premises infrastructure, monitoring, alerting, access controls, and platform integrations * Drive automation, Infrastructure as Code, Ansible, configuration management, and repeatable operational patterns that reduce toil and improve reliability * Guide major upgrades, migrations, lifecycle planning, patch strategies, recovery planning, and technical roadmaps for supported platforms * Lead the most complex incidents and technical investigations, including root cause analysis, corrective action planning, and systemic reliability improvements * Mentor and technically guide SREs in operational excellence, troubleshooting, automation, secure administration, and architectural thinking * Partner with program leadership, cybersecurity, infrastructure, software engineering, database, networking, suppliers, customers, and other stakeholders * Identify platform risks, technical debt, capacity constraints, single points of failure, compliance concerns, and operational gaps, then drive remediation plans * Lead efforts to operationally field higher-quality end-to-end system software more frequently * Participate in after-hours support and escalation for urgent or mission-impacting issues as required, DevOps Gitlab Ansible Tooling Jenkins Dashboard SonarQube DevSecOps Operations Automation PostgreSQL Code Review Artifactory End Systems Registration GitLab CI/CD Investigation Cyber Security Backup Devices Building Codes Export Control Version Control Microsoft Azure Access Controls Cloud Computing Customer Service SAP Applications Data Maintenance Secret Clearance Patch Management Health Insurance Operating Systems Change Management Incident Response CompTIA Security+ Disaster Recovery Windows PowerShell Security Clearance Workflow Management Root Cause Analysis Process Improvement Amazon Web Services Incident Management Pull/Merge Requests Software Deployment Software Development Atlassian Confluence System Administration Collective Bargaining Disaster Recovery Plan Database Administration Configuration Management Bash (Scripting Language) C# (Programming Language) SQL (Programming Language) Agile Software Development C++ (Programming Language) Backup and Recovery System Java (Programming Language) Verbal Communication Skills Database Storage Structures Data-Driven Decision Making Standard Operating Procedure Git (Version Control System) Site Reliability Engineering Infrastructure as Code (IaC) Python (Programming Language) Technical Delivery Management Security Requirements Analysis Key Performance Indicators (KPIs) Information Technology Operations Troubleshooting (Problem Solving) +0 Google IT Automation with Python ## Related Videos - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) - [GitOps for the people](https://www.wearedevelopers.com/videos/461-gitops-for-the-people) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [GitLab CI pipelines for a whole company](https://www.wearedevelopers.com/videos/143-gitlab-ci-pipelines-for-a-whole-company) - [Building a Multi-Agent Orchestration Engine That Actually Follows the Rules](https://www.wearedevelopers.com/videos/100159-building-a-multi-agent-orchestration-engine-that-actually-follows-the-rules) - [Improving quality with Agentic AI with Rovo Dev and Xray](https://www.wearedevelopers.com/videos/2005-improving-quality-with-agentic-ai-with-rovo-dev-and-xray) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers) - [Dev Digest 134 - Where pixels sing?](https://www.wearedevelopers.com/magazine/477-dev-digest-134-where-pixels-sing) - [Now is the time for industrialized software development](https://www.wearedevelopers.com/magazine/601-now-is-the-time-for-industrialized-software-development) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)