> Markdown version of [/jobs/ext/1993663-site-reliability-engineer-in-rockville](https://www.wearedevelopers.com/jobs/ext/1993663-site-reliability-engineer-in-rockville). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer in Rockville - **Company:** Energy Jobline - **Location:** Rockville, MD, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Cloud Engineering, Continuous Integration, Data Systems, Disaster Recovery, Distributed Systems, Python (Programming Language), Reliability Engineering, Ansible, Large Language Models, Cloudformation, Containerization, Gitlab-ci, Infrastructure Automation Frameworks, Information Technology, Cloudwatch, Terraform, Splunk, New Relic (SaaS), Docker, Jenkins, Vulnerability Analysis - **Published:** August 8, 2026 - **Apply:** https://www.energyjobline.com/job/site-reliability-engineer-rockville-31387194 ## About the Role Do you have a real feel for how distributed systems behave, and a knack for tracking down the network, infrastructure, or pipeline issue everyone else gave up on? Are you comfortable in the cloud, fluent in CI/CD, and the type who believes an alert should mean something and a dashboard should tell a story? If you love keeping complex systems healthy, fast, and quietly reliable, then apply. Like, now. Come join us if you're motivated to learn from others, to learn from mistakes, to be part of a future-looking and growth-oriented team., * A bachelor's degree in computer science, engineering, or a related field (or equivalent hands-on experience). * 3-5 years of experience in site reliability, systems, or cloud engineering, with meaningful time spent in AWS environments. * Solid working knowledge of core AWS services, architecture, and best practices. * Hands-on experience with infrastructure-as-code tools (Terraform, Ansible, or CloudFormation). * A good understanding of CI/CD pipelines and automation tools (Jenkins, GitLab CI, or similar). * Comfort scripting and automating in Python. * Familiarity with monitoring and observability tooling (CloudWatch, New Relic, Splunk, or comparable). * Strong problem-solving instincts and the composure to work calmly under pressure. * Clear communication skills, with the ability to make complex technical concepts understandable. What would blow us away: * You've previously worked with CMS. * You have experience working in AI, NLP, or LLM-driven environments. * You have all the AWS certifications and the real-world scars that come with them. ## Description Job DescriptionJob DescriptionWe are Skyward. That is, a love for people, for improvement, for human advancement through information technology. We are a people-centered business with a desire to serve others. We are diverse and unified; creative and collaborative; a collection of complementary, not competing talents. And though on the surface we remain relaxed, beneath, a torrent of energy links us to our civic tech mission. We stand by our values, and we won't compromise on any of them. Integrity: We're conscientious, intentional, and empathetic. Our words and actions align. That's our character. Please don't ask us to play another part, we're poor actors. Compassionate: If we may borrow a quote from Theodore Roosevelt: "No one cares how much you know until they know how much you care." Because our team is thoughtful and supportive, caring deeply for each other, our clients, and our work, this comes naturally. Inquisitive: We remain students by failing openly and turning lessons into solutions.Unconventional: For us, life isn't what happens outside of work. Work happens inside of life and our culture erases the line often dividing the two. Authentic: Made possible only because we embody the values listed above. We're relaxed and fun yet intensely curious and driven. Team members are placed with thought, care, and precision to ensure that Trust, Truth, and Transparency continue to represent our brand. Because of that, we continue Onward, Upward, and Skyward., * Join the team supporting the Centers for Medicare & Medicaid Services (CMS) as it merges and modernizes its enterprise knowledge and data systems into a single, AI-driven platform, reducing manual effort, improving data accuracy, and enhancing transparency for stakeholders. * Keep the systems up and the users happy. Operate and tune AWS environments to meet infrastructure and application availability SLAs, even during transition and change. * Build observability that actually informs. Implement continuous monitoring, alerting, and dashboards using tools like AWS CloudWatch, New Relic, and Splunk, and establish performance baselines so you can spot degradation before users do. * Automate the toil. Write infrastructure-as-code (Terraform, Ansible) and support CI/CD pipelines (Jenkins) and containerized workloads (Docker) for repeatable, reliable deployments. * Define and track the numbers that matter. Set and monitor SLIs and SLOs, and produce performance, load/stress, and bottleneck reports that drive smarter decisions. * Optimize for performance, security, and cost. Use tools like AWS Trusted Advisor to find and act on improvement opportunities. * Support security and compliance modernization. Partner with the Security & Compliance SME to review vulnerability and security scans, feed continuous monitoring, and help advance the move toward a Continuous ATO (cATO) within a FISMA Moderate boundary (RMF, ARS, IS2P2). * Strengthen resilience. Help design and maintain disaster recovery and COOP continuity so the systems hold up against outages, incidents, and the unexpected. * Own incidents end to end. Drive response, run blameless post-mortems, and implement the preventative fixes that keep the same thing from happening twice. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [The Road to MLOps: How Verivox Transitioned to AWS](https://www.wearedevelopers.com/videos/1050-the-road-to-mlops-how-verivox-transitioned-to-aws) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [How we built an AI-powered code reviewer in 80 hours](https://www.wearedevelopers.com/videos/1511-how-we-built-an-ai-powered-code-reviewer-in-80-hours) - [Our GitOps approach for deploying an Identity Provider and an API Gateway in a SaaS company](https://www.wearedevelopers.com/videos/776-our-gitops-approach-for-deploying-an-identity-provider-and-an-api-gateway-in-a-saas-company) ## Related Articles - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this)