> Markdown version of [/jobs/ext/3530692-site-reliability-engineer-fedramp](https://www.wearedevelopers.com/jobs/ext/3530692-site-reliability-engineer-fedramp). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer- FedRamp - **Company:** Cisco Systems, Inc. - **Location:** Boxborough, MA, United States - **Experience:** Expert - **Salary:** $128,600.0 - $184,900.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Amazon Web Services, Audit Trail, Bash Shell, Cloud Computing, Cloud Engineering, Continuous Integration, DevOps, Disaster Recovery, Distributed Systems, Elasticsearch, Identity and Access Management, Python (Programming Language), Network Troubleshooting, Linux System Administration, Reliability Engineering, Prometheus, Runbook, Software Deployment, Software Engineering, Software Vulnerability Management, Cisco WebEx, Data Logging, Cloud Platform System, Grafana, Software Troubleshooting, Gitlab, Git, Kubernetes, Information Technology, Maintaining Code, Cloudwatch, Terraform, Splunk, Docker, Jenkins, Golang, Programming Languages, Microservices - **Published:** October 2, 2026 - **Apply:** https://dejobs.org/x/x/356D20F0C7E641C1BCF9F450DD15A613/job/ ## About the Role * Bachelor's degree in Computer Science, engineering, or related field, plus 5+ years of related experience, or equivalent practical experience. * 5+ years of Software development or automation experience using Java, Go, Python, or a comparable programming language, including testing, troubleshooting, and maintaining code. * Experience supporting cloud infrastructure, production services, DevOps, platform engineering, or Site Reliability Engineering in AWS. * Experience with Linux administration and troubleshooting application, system, or networking issues in production or production-like environments. * Experience with Kubernetes, Docker, microservices, and cloud-native application deployment/troubleshooting, including Git, CI/CD pipelines, and infrastructure-as-code/automation (e.g., Terraform, Python, Bash, or Go)., * Experience operating services in a FedRAMP, government cloud, or other regulated environment. * Experience using logs, metrics, dashboards, and alerts to troubleshoot production services with tools such as Prometheus, Grafana, CloudWatch, CloudTrail, Elastic Stack, or Splunk. * Experience supporting CI/CD pipelines and automated production deployments using tools such as GitLab or Jenkins. * Experience with highly available, multi-region distributed systems, capacity planning, and disaster recovery. * Strong written and verbal communication skills with experience creating runbooks, troubleshooting guides, operational procedures, and audit documentation. * Experience participating in an on-call rotation, production incidents, root-cause analysis, or post-incident reviews, with an understanding of Incident Commander responsibilities. ## Description This role will deploy, operate, and maintain resilient AWS and Kubernetes-based infrastructure and microservices that support the Webex for Government environment. You will monitor service health, troubleshoot complex incidents, and help restore service quickly in a 24x7 production environment. You will automate repeatable operational work and infrastructure provisioning to improve safety, efficiency, and consistency across deployments. You will strengthen platform security and compliance by supporting identity and access management, logging, hardening, vulnerability remediation, and FedRAMP continuous monitoring activities. You will collaborate across engineering, security, and compliance teams to identify operational risks, improve reliability, and build documentation, runbooks, and incident reports that make the environment easier to operate. ## Related Videos - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [GitLab CI pipelines for a whole company](https://www.wearedevelopers.com/videos/143-gitlab-ci-pipelines-for-a-whole-company) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) - [WeAreDevelopers LIVE - Modern DevOps for IoT Devices and More](https://www.wearedevelopers.com/videos/1805-wearedevelopers-live-modern-devops-for-iot-devices-and-more) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [The Best Software Developer Blogs to Read](https://www.wearedevelopers.com/magazine/156-the-best-software-developer-blogs-to-read) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Best Countries for Software Engineers](https://www.wearedevelopers.com/magazine/267-best-countries-for-software-engineers)