Site Reliability Engineer

Booz Allen Hamilton Inc.
Arlington, VA, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Part-time / full-time
Experience level
Expert
Experience required
5 years minimum
Compensation
$86,800.0 - $198,000.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Amazon Elastic Compute Cloud Bash Shell Configuration Management Linux DevOps Disaster Recovery Github Python (Programming Language) Reliability Engineering Ansible Prometheus
+17 more
Scripting System Availability Grafana Cloudformation SC Clearance Gitlab-ci Kubernetes Cloud Optimization Cloudwatch Puppet Terraform Splunk Dynatrace Serverless Computing Jenkins Vulnerability Analysis Programming Languages

Requirements

  • 5+ years of experience in an SRE, DevOps, or production engineering role with direct ownership of production reliability and uptime
  • Experience with AWS production across EC2-based, containerized, and serverless workloads
  • Experience with container orchestration in production, such as Kubernetes, EKS, or ECS, including deployment, scaling, and troubleshooting
  • Experience with Infrastructure as Code, such as Terraform or CloudFormation, for provisioning and managing cloud resources and building and maintaining CI/CD pipelines, such as GitLab CI, GitHub Actions, or Jenkins
  • Experience with observability and monitoring with tools, such as Grafana, Prometheus, Splunk, and CloudWatch, including building dashboards and alerting
  • Experience in at least one scripting or programming language, such as Python, Go, or Bash, for automation and tooling, and Linux and Windows administration and troubleshooting in production environments
  • Experience with incident response, root cause analysis, and defining SLAs, SLOs, and error budgets
  • Knowledge of vulnerability scanning and remediation as part of the deployment lifecycle
  • Ability to obtain a Secret clearance
  • Bachelor’s degree

Nice If You Have:

  • Experience designing for high availability, multi-AZ/multi-region, and disaster recovery
  • Experience writing blameless postmortems or after action reports (AAR) deriving root cause, corrective actions and driving reliability improvements across teams so it never happens again
  • Experience with distributed tracing and OpenTelemetry instrumentation
  • Experience with configuration management tooling, such as Ansible, Chef, or Puppet
  • Experience with cloud cost optimization and FinOps practices
  • Experience in regulated or compliance-driven environments, such as FedRAMP, NIST 800-53, or SOC 1
  • AWS certification, such as Solutions Architect, SysOps, or DevOps Engineer

Benefits & conditions

At Booz Allen, we celebrate your contributions, provide you with opportunities and choices, and support your total well-being. Our offerings include health, life, disability, financial, and retirement benefits, as well as paid leave, professional development, tuition assistance, work-life programs, and dependent care. Our recognition awards program acknowledges employees for exceptional performance and superior demonstration of our values. Full-time and part-time employees working at least 20 hours a week on a regular basis are eligible to participate in Booz Allen’s benefit programs. Individuals that do not meet the threshold are only eligible for select offerings, not inclusive of health benefits. We encourage you to learn more about our total benefits by visiting the Resource page on our Careers site and reviewing Our Employee Benefits page.

Salary at Booz Allen is determined by various factors, including but not limited to location, the individual’s particular combination of education, knowledge, skills, competencies, and experience, as well as contract-specific affordability and organizational requirements. The projected compensation range for this position is $86,800.00 to $198,000.00 (annualized USD). The estimate displayed represents the typical salary range for this position and is just one component of Booz Allen’s total compensation package for employees. This posting will close within 90 days from the Posting Date.

Identity Statement

As part of the hiring process, we will ask you to complete an identity verification process that leverages advanced biometrics and artificial intelligence to ensure authenticity and protect against identity fraud. You are expected to be on camera during interviews and assessments. We reserve the right to take your picture to verify your identity and prevent fraud.

About the company

AI is a part of our daily work at Booz Allen, and we are committed to the responsible and ethical use of AI tools. However, we want to ensure a fair candidate process based on your own skills and knowledge. As part of this commitment, the use of artificial intelligence (AI) or other tools to assist with responses during interviews (whether in-person or virtual) is prohibited unless permission is explicitly provided.

Work Model

Our people-first culture prioritizes the benefits of collaboration, whether it occurs in person or virtually. To support engagement and effective communication, employees working virtually are generally expected to have their cameras on during meetings.

  • Remote: If this position is listed as remote, there may still be occasions when you are required to work in person at a Booz Allen or customer facility.
  • Hybrid: If this position is listed as hybrid, you will be expected to work from a Booz Allen facility frequently, in alignment with leadership expectations and the needs of the role. You may also be required to work from or visit a customer facility.
  • Onsite: If this position is listed as onsite, work will primarily be performed at a Booz Allen office or customer facility, where employees will collaborate directly with colleagues and customers as required by the role.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.clearancejobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:26 min

Understanding Puppeteer and its underlying architectural design

Miki Lombardi · JS Congress

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · WWC 2023

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

4:01 min

Comparing Terraform to popular configuration management tools

Devlin Duldulao · LIVE

Videos

See all

Related articles

See all