> Markdown version of [/jobs/ext/2915495-site-reliability-engineer-cyber](https://www.wearedevelopers.com/jobs/ext/2915495-site-reliability-engineer-cyber). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer, Cyber - **Company:** Anduril Industries - **Location:** Arlington, VA, United States - **Experience:** Experienced - **Salary:** $146,000.0 - $220,000.0 - **Contract:** Permanent contract - **Skills:** Automation of Tests, Bash Shell, Continuous Integration, Software Debugging, Linux, DevOps, Distributed Systems, Firmware, Hardware-In-The-Loop Simulation, Python (Programming Language), Network Troubleshooting, Networking Basics, Reliability Engineering, Software Deployment, Software Engineering, Software Systems, Systems Integration, Kubernetes - **Published:** September 15, 2026 - **Apply:** https://www.clearancejobs.com/jobs/9163360/site-reliability-engineer-cyber ## About the Role * Currently possesses and is able to maintain an active U.S. TS/SCI security clearance. * Based in the DC metro area to support 3-5 days per week working on site at customer facilities. * 4+ years of experience in a Sys Admin, Site Reliability, DevOps, or Software Engineering role. * Deep, practical experience with Linux and Kubernetes (or a similar container orchestrator). * Working knowledge of network fundamentals and the ability to debug connectivity in a locked-down environment. * Experience delivering and maintaining systems on air-gapped and security-hardened networks. * Strong proficiency in Python or Bash for automation and debugging, and the ability to read and debug service code in a compiled language such as Go. * Excellent written and verbal communication skills for collaborating with a cross-functional engineering team and external customers., * Experience in debugging and resolving networking issues. * Ability to quickly understand and navigate complex, multi-disciplinary systems and established codebases. * Experience building automation for hardware-in-the-loop (HIL) or software-in-the-loop (SIL) test environments. * Ability to drive consensus across internal and external stakeholders. ## Description As a Site Reliability Engineer in Anduril Cyber, you will solve a wide variety of problems involving networking, systems integration, distributed systems, and more, while making pragmatic engineering tradeoffs along the way. Your efforts will ensure that Anduril's software is reliable, scalable, and deployable, in order to achieve critical national security outcomes. You will work closely with software developers, customers, and external vendors to get working offensive Cyber products into the hands of customers. You will own the full deployment pipeline - from CI/CD and pre-production test environments, through canary deployments in customer-hosted integration environments, to production in air-gapped enclaves., You will also be the steward of Anduril's mission and technological advantage in the room with customers - attending technical meetings, explaining why the system behaves the way it does, and absorbing requirements firsthand. What you observe on site is the primary input to our team's roadmap. Site Reliability Engineers must be driven by a "Whatever It Takes" mindset - executing in an expedient, scalable, and pragmatic way while keeping the mission top-of-mind and making sound decisions to deliver successful outcomes on-time and with high quality. WHAT YOU'LL DO * Own the health of our deployed systems and keep them running with minimal downtime. * Automate and improve our software deployment processes into air-gapped, TS/SCI environments. * Design, build, and maintain the CI/CD and automated test infrastructure for Cyber's complex hardware and software systems. * Develop metrics dashboards, TUIs, scripts, and other tools that automate common deployment steps or help to debug our software stack. * Drive engineering requirements based on onsite observations. * Perform root cause analysis and diagnose issues in mission-critical systems across our software stack, the Lattice OS stack, and external vendor services. * Build strong relationships with internal and external customers to identify technical solutions to their problems. * Drive continuous improvement by instrumenting systems, analyzing failures, and leading post-mortem events that span software, firmware, and hardware., To ensure your safety and help you navigate your job search with confidence, please keep the following critical points in mind: * No Financial Requests: Anduril will never solicit payment or demand personal financial details (such as banking information, credit card numbers, or social security numbers) at any stage of our hiring process. Our legitimate recruitment is entirely free for candidates. ## Related Videos - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) - [Playing Pong on a shoulder press machine](https://www.wearedevelopers.com/videos/100140-playing-pong-on-a-shoulder-press-machine) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Agent Smith Gets Hardware: Autonomous IoT Hacking From Debug Port to Cloud API](https://www.wearedevelopers.com/videos/100258-agent-smith-gets-hardware-autonomous-iot-hacking-from-debug-port-to-cloud-api) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 134 - Where pixels sing?](https://www.wearedevelopers.com/magazine/477-dev-digest-134-where-pixels-sing) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this)