> Markdown version of [/jobs/ext/2105882-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2105882-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Anduril Industries - **Location:** United States - **Experience:** Expert - **Salary:** $166,000.0 - $250,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Confluence, JIRA, Microsoft Azure, C++ (Programming Language), Collaborative Software, Nvidia CUDA, Continuous Integration, DevOps, Elasticsearch, Github, Monitoring of Systems, Python (Programming Language), Machine Learning, Open Source Technology, OpenCL, Reliability Engineering, Logstash, Ansible, Prometheus, Software Deployment, Rust (Programming Language), Circleci, Data Logging, Google Cloud, Grafana, Parallel Computation, Infrastructure as Code (IaC), Git, Containerization, Kubernetes, Infrastructure Automation Frameworks, Build Tools, Machine Learning Operations, Kibana, Terraform, Serverless Computing, Docker, Elk Stack, Artifactory - **Published:** August 18, 2026 - **Apply:** https://www.dice.com/job-detail/c06c4ca0-953d-414c-8ac3-c012100d6160 ## About the Role * Advanced proficiency in programming languages (Python for scripting and integration). * Experience with CI/CD tools like GitHub Actions, Jfrog Artifactory, Git, and CircleCI. * Proficiency with IaC tools (Terraform, Ansible). * Experience with cloud platforms (Azure, AWS, Google Cloud Platform). * Proficiency in containerization (Docker) and container orchestration (Kubernetes). * Knowledge of model registries and feature stores (e.g., MLflow, Kubeflow). * Experience with logging and monitoring tools ( Prometheus, Grafana). * Understanding of parallel computing frameworks (CUDA, OpenCL). * Strong collaboration skills and proficiency with collaborative tools (JIRA, Confluence). * Eligible to obtain and maintain an active U.S. Secret security clearance., * Experience writing cloud-native services using C++, Rust, Python and/or Go * Familiarity with observability concepts and tools. * Knowledge of security best practices for DevOps and MLOps. ## Description Anduril Maritime delivers platforms, systems, and integrated effects in the maritime domain. Our autonomous vehicles (sub-surface and surface) are the cornerstone of these capabilities, and we continually strive to push the boundaries of the possible in terms of endurance, autonomy and mission capability. The Maritime team develops and maintains core products and payloads, and adapts and applies those products to serve a wide variety of defense, IC and commercial customers in US and international markets. ABOUT THE JOB As a Senior Site Reliability Engineer on the Undersea Dominance team, you will build and operate the infrastructure that keeps our operational and production systems running at full speed. You'll develop and manage CI/CD pipelines, automate infrastructure with code, and deploy applications across cloud and edge environments with security, traceability, and reliability in mind. You'll work closely with software, data, and operations engineers to turn designs into working systems-streamlining development, improving performance, and keeping production stable as we scale. You'll also collaborate with digital, manufacturing, and corporate technology teams across Anduril in a high-tech, fast-paced culture of innovation focused on solving real problems and delivering results. If you're driven to build systems that last, thrive on deep technical challenges, and want to see your work directly shape how we design, build, and sustain complex platforms, you'll be helping build the future of digital shipbuilding and the next generation of maritime vehicles. WHAT YOU'LL DO * Build and Manage CI/CD Pipelines: Develop and maintain CI/CD pipelines using tools like GitHub Actions and Jfrog Artifactory to ensure seamless integration and deployment of machine learning models and applications. * Infrastructure as Code (IaC): Utilize Terraform and Ansible to automate infrastructure provisioning and management on cloud platforms such as Azure, AWS, or Google Cloud Platform (Google Cloud Platform). * Containerization and Orchestration: Implement containerization solutions with Docker and manage container orchestration using Kubernetes to ensure reliable deployment and scaling of applications. * Monitoring and Logging: Establish comprehensive monitoring and logging solutions using tools like ELK Stack (Elasticsearch, Logstash, Kibana), Prometheus, and Grafana to ensure the smooth operation of deployment environments. * Collaborate with Cross-Functional Teams: Work closely with development, data science, and operations teams to foster collaboration and ensure the efficient and effective deployment of machine learning models., To ensure your safety and help you navigate your job search with confidence, please keep the following critical points in mind: * No Financial Requests: Anduril will never solicit payment or demand personal financial details (such as banking information, credit card numbers, or social security numbers) at any stage of our hiring process. Our legitimate recruitment is entirely free for candidates. * Please always verify communications: + Direct from Anduril: If you receive an email from one of our recruiters, it will only come from an @anduril.com address. + Via Agency Partner: If contacted by a recruiting agency for an Anduril role, their email will clearly identify their agency. If you suspect any suspicious activity, please verify the agency's authenticity by reaching out to * Exercise Caution with Unsolicited Outreach: If you receive any communication that appears suspicious, contains grammatical errors, or makes unusual requests, do not engage. Always confirm the sender's email domain is @anduril.com before providing any personal information or clicking on links. * What to Do If You Suspect Fraud: Should you encounter any questionable or fraudulent outreach claiming to be from Anduril, please report it immediately to Your proactive caution is invaluable in protecting your personal information and upholding the security and trustworthiness of our recruitment efforts. Data Privacy To view Anduril's candidate data privacy policy, please visit By submitting your application, you consent to Anduril Industries using a third-party service provider to conduct pre-employment risk, integrity, and due diligence screening and assessing potential risks as part of your application process. This third-party service provider provides risk-intelligence services that may include analysis of sanctions and watchlists, adverse media, public-record information, and other lawful open-source or commercial data sources. This third-party service provider does not act as a consumer reporting agency. Use of this provider helps to ensure compliance with applicable laws and protect technology, intellectual property, and organizational security. ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Improving quality with Agentic AI with Rovo Dev and Xray](https://www.wearedevelopers.com/videos/2005-improving-quality-with-agentic-ai-with-rovo-dev-and-xray) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) - [Designing UX for SRE Agents in High-Stakes Incidents](https://www.wearedevelopers.com/videos/100003-designing-ux-for-sre-agents-in-high-stakes-incidents) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix)