Cloud Reliability Engineer

TEKNOLUXION CONSULTING, LLC
Springfield, VA, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Systems Engineering Cloud Computing DevOps IT Management Reliability Engineering Cloud Services Software Engineering Devsecops

Job description

At Bcore, our strength comes from how we deliver impact to the mission. Whether it’s architecting critical IT solutions, producing actionable intelligence, or developing cutting edge technology, we succeed because of the expertise, collaboration, and agility of our teams. Our Mission Services division combines enterprise IT, cloud solutions, DevSecOps, systems engineering, software development, and operational support. Bcore accelerates decisive advantage for warfighters and intelligence professionals by fusing human insight, rapid-fire engineering, precision-measured outcomes, and relentless grit into mission-ready solutions. Do you want to join a team that is building tailored technical solutions to modernize our government’s mission and our client’s business? Do you have a desire to change how people work? Are you interested in helping to protect our nation’s cyber interests? Join our growing team supporting the NGA customer missions as an Cloud Reliability Engineer. Responsibilities: What you get to do every day:

  • Deliver resilient architecture designs (multi-AZ and/or multi-Region) tailored to customer workloads
  • Conduct Well-Architected Reviews focused on the Reliability Pilla
  • Implement and validate resilience through chaos engineering and game days
  • Establish operational runbooks, monitoring, and DR procedures

Requirements

Clearance Required: Active TS/SCI with CI poly, * Conducting Well-Architected Reviews with a Reliability Pillar focus

  • Workshop facilitation: resilience assessments, architecture reviews, and game days
  • Stakeholder communication - translating technical resilience concepts to business value
  • Documentation: architecture decision records (ADRs), resilience runbooks, SOPs
  • Customer relationship management and executive-level reporting
  • Cross-functional collaboration with customer engineering, security, and operations teams
  • Mentoring customer teams on resilience best practices and operational maturity

What is ideal?

  • AWS Resilience Hub hands-on experience (resilience assessment and policy management)
  • AWS Certifications: Solutions Architect Professional, DevOps Engineer Professional
  • Experience in regulated industries (FedRAMP, IL4/IL5 environments)
  • Familiarity with NIST SP 800-160 (Systems Security Engineering) or DoD Mission Assurance frameworks
  • Experience with Service Control Policies (SCPs) and multi-account resilience governance
  • Experience with Orange and Purple Cloud PMO capabilities and Authorization Frameworks
  • Prior ProServe or consulting delivery experience with large enterprise customers
  • Hands-on experience with third-party resilience tooling (Gremlin, LitmusChaos)
  • Intelligence Community Experience preferred

Benefits & conditions

Pulled from the full job description

  • Referral program
  • 401(k)
  • Health insurance
  • Paid time off
  • Vision insurance
  • Dental insurance
  • Life insurance, * The expected salary range within the Washington, DC metropolitan area is: INSERT-THE-RANGE. Final compensation is unique to each individual and will be determined based on factors such as experience, education, geographic location, and contractual requirements. This is not a guarantee.
  • Benefits include Health/Dental/Vision, 401(k), Paid Time Off, STD/LTD/Life Insurance/Voluntary Life Insurance, Stipends, Referral Bonuses, and more.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

1:15 min

Key lessons learned from implementing automated mobile DevSecOps

Moataz Nabil Moataz Nabil · LIVE

1:53 min

Managing infrastructure limitations with managed Amazon Aurora databases

Dharin Shah Dharin Shah · WWC 2025

1:37 min

Defining reliability through the AWS well-architected framework

Florian Mair Florian Mair · WWC 2024

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

1:24 min

Evaluating formal AWS certifications versus raw practical engineering experience

Jan Giacomelli · LIVE

Videos

See all

Related articles

See all