Site Reliability Engineer (SRE)

Quindar Inc.
Washington, DC, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$160,000.0 - $200,000.0
Working hours
Shift work
Job source

Tech stack

Application Programming Interfaces (APIs) Amazon Web Services Cloud Computing Security Databases Continuous Integration DevOps Distributed Systems Fault Tolerance Identity and Access Management Virtual Private Networks (VPN) Python (Programming Language) Networking Basics
+24 more
Reliability Engineering Datadog Transport Layer Security Load Balancing Cloud Platform System Okta Grafana Caching Reliability of Systems Backend Gitlab Git Event Driven Architecture Containerization Kubernetes Infrastructure Automation Frameworks Information Technology Deployment Automation Rancher Front End Software Development Terraform Virtual Private Clouds Devsecops AWS EKS

Job description

Design, automate, deploy, and operate highly reliable cloud systems supporting mission-critical workloads for U.S. Government customers. This role is centered on DevSecOps and site reliability engineering, with a strong emphasis on deployment automation, operational stability, and system resilience across AWS GovCloud and AWS C2E environments.

You will be responsible for continuously improving the reliability and operability of Quindar’s platform in production, ensuring systems are observable, fault-tolerant, and require minimal manual intervention. Your work will directly impact mission success by improving system uptime, deployment velocity, and operational confidence in constrained and classified environments.

A key focus of this role is building and evolving automated deployment pipelines, hardened runtime environments, and repeatable infrastructure patterns that support secure and scalable operations in regulated environments.

You will also support and improve Quindar deployments to air-gapped networks, driving consistency, reliability, and performance across all environments. As the organization grows, you will help define and implement best practices for availability, latency, incident response, and service-level objectives (SLOs).

This role includes participation in incident response and a 24/7 on-call rotation, with a strong mandate to eliminate toil through automation and continuously improve system reliability.

You will collaborate closely with frontend, backend, and platform engineers to ensure systems meet performance, reliability, and mission assurance requirements.

Requirements

Do you have experience in Virtual Private Clouds?, Do you have a Bachelor’s degree?, * Strong experience with Kubernetes and containerized workloads in production environments

  • Hands-on experience operating clusters in AWS EKS, Rancher, or similar platforms
  • Experience supporting GovCloud, IL-enclave, or C2E environments
  • Deep experience with CI/CD systems and deployment automation (GitLab preferred)
  • Proficiency in Python and Infrastructure-as-Code tools (Terraform or similar)
  • Experience with observability platforms (Grafana LGTM stack, Datadog, or equivalent)
  • Strong understanding of distributed systems, APIs, databases, caching, and event-driven architectures
  • Solid networking fundamentals (VPCs, VPNs, load balancers, TLS, service connectivity)
  • Experience with Linux/Unix systems
  • Familiarity with cloud security best practices, enclave boundaries, and secure system design
  • Experience with identity and access management (AWS IAM, Auth0, Keycloak, ICAM patterns)
  • Strong Git fundamentals and experience supporting deployments across multiple classification levels, * Bachelor’s degree in Computer Science or related field
  • 3+ years of professional experience as an SRE, DevOps, reliability, infrastructure, or platform engineer
  • Active U.S. Security Clearance preffered
  • Experience working toward ATO/authorization in federal, DoD, or IC environments preferred
  • Experience supporting deployments in GovCloud, C2S/C2E, or IL-enclave environments highly desirable, * To conform to U.S. Government export regulations, applicant must be a (i) U.S. citizen or national, (ii) U.S. lawful, permanent resident (aka green card holder), (iii) Refugee under 8 U.S.C. § 1157, or (iv) Asylee under 8 U.S.C. § 1158, or be eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here.

Benefits & conditions

Pulled from the full job description

  • 401(k) 4% Match
  • 401(k) matching
  • Unlimited paid time off, * We work in a cutting edge industry and you will get the opportunity to be part of a small team with a large direct impact on the success of our customers’ space missions!
  • We take work life balance very seriously. We require employees to take 15 days off but provide unlimited PTO and follow most US federal government holidays.
  • Mental health is just as important as physical so we provide quarterly health & wellness benefits.
  • Comprehensive health insurance for you and your family with 100% coverage for employees.
  • We encourage employees to save for retirement and provide 4% 401(k) matching.
  • Annually we have a 4-day company offsite. Previous locations include San Francisco, Nashville, Denver, Santa Fe, New Orleans, San Diego, Bozeman, and New York City.
  • Our culture and company is evolving. You will be key in creating the next major or minor version!

Compensation Range: $160K - $200K

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

6:14 min

Structuring CI/CD pipelines with integrated security and quality checks

Christoph Ruggenthaler · LIVE

2:33 min

Introduction to security advocacy and automation testing

Chris Heilmann +2 · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

4:54 min

Implementing geographic salary tiers for compensation equity and fairness

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

4:37 min

Architecting single sign-on flows across multiple application domains

Gift Egwuenu · WWC 2023

Videos

See all

Related articles

See all