> Markdown version of [/jobs/ext/2734630-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2734630-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Zachary Piper - **Location:** United States (Remote available) - **Experience:** Experienced - **Salary:** $140,000.0 - $165,000.0 - **Contract:** Permanent contract - **Skills:** Kubernetes Security, Amazon Web Services, Microsoft Azure, Cloud Computing, Configuration Management, Computer Programming, Continuous Integration, Linux, DevOps, Monitoring of Systems, Python (Programming Language), Linux System Administration, Octopus Deploy, Object-Oriented Software Development, Performance Tuning, Reliability Engineering, Prometheus, Ruby, Scripting, System Availability, Grafana, Infrastructure as Code (IaC), Cloudformation, Containerization, Git Flow, Kubernetes, Infrastructure Automation Frameworks, Deployment Automation, Terraform, Docker, Golang - **Published:** September 5, 2026 - **Apply:** https://www.clearancejobs.com/jobs/9145957/site-reliability-engineer ## About the Role · 4-6 years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or related infrastructure-focused roles. · Strong hands-on experience managing and supporting Kubernetes clusters in production environments. · Experience with AWS, Azure, or similar cloud platforms; GovCloud experience is a plus. · Strong Linux administration skills, including system troubleshooting, performance analysis, and infrastructure support. · Experience with Infrastructure as Code technologies such as Terraform, CloudFormation, or similar tools. · Programming or scripting experience in Python, Go, Ruby, or other object-oriented languages.. · Experience with observability and monitoring tools such as Prometheus, Grafana, and centralized logging solutions. · Exposure to CI/CD tools, deployment automation, and GitOps practices such as ArgoCD is preferred. · Experience supporting FedRAMP High, DoD IL5, regulated, or audited environments is strongly preferred., Keywords: Kubernetes, Site Reliability Engineering (SRE), DevOps, Platform Engineering, AWS, Azure, GovCloud, Linux, Terraform, CloudFormation, Infrastructure as Code (IaC), CI/CD, ArgoCD, Prometheus, Grafana, Monitoring, Alerting, Observability, Containerization, Docker, Networking, Automation, Python, Go, Ruby, Object-Oriented Programming, Cloud Infrastructure, Kubernetes Clusters, Production Support, Troubleshooting, Incident Response, Performance Optimization, Scalability, Reliability, High Availability, Security, Compliance, FedRAMP High, DoD IL5, Continuous Monitoring, Audit Support, GitOps, Deployment Automation, Platform Operations, Container Security, Cross-Functional Collaboration, SLI, SLO, Error Budgets, On-Call Support, Systems Administration, Infrastructure Management, Cloud-Native Technologies, AWS Infrastructure, Infrastructure Automation, Root Cause Analysis, Configuration Management, Platform Reliability, Regulated Environments. ## Description Piper Companies is seeking a Site Reliability Engineer (SRE) to support the development, maintenance, and operation of a Kubernetes-based platform within highly regulated cloud and on-premises environments. This individual will work closely with senior engineers and technical leaders to improve platform reliability, scalability, security, and operational performance while supporting compliance-driven initiatives. This is a long-tern contract opportunity with a strong focus on Kubernetes infrastructure, Linux administration, automation, and cloud technologies. This individual may sit remote in the US., · Build, maintain, and support Kubernetes clusters across on-premises and AWS environments. · Monitor platform reliability, availability, and performance through observability, alerting, and troubleshooting activities. · Develop and implement automation tools and processes to improve operational efficiency and reduce manual intervention. · Collaborate with senior engineers to define and track service reliability metrics, including SLIs, SLOs, and error budgets. · Support compliance, security, auditing, and continuous monitoring initiatives within regulated environments. · Contribute to Infrastructure as Code (IaC) development and enhancements using tools such as Terraform, CloudFormation, or similar technologies. · Improve CI/CD pipelines and deployment processes to support platform scalability and operational excellence. · Partner with Security, Platform, and Application teams to resolve issues and deliver reliable infrastructure solutions. ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Learning Kubernetes made easy with KubeCampus](https://www.wearedevelopers.com/magazine/348-learning-kubernetes-made-easy-with-kubecampus) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers)