> Markdown version of [/jobs/ext/739306-site-reliability-engineer-2](https://www.wearedevelopers.com/jobs/ext/739306-site-reliability-engineer-2). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer 2 - **Company:** Chags Health Information Technology LLC - **Location:** Columbia, MD, United States (Remote available) - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Continuous Integration, DevOps, Disaster Recovery, Github, Monitoring of Systems, HP Systems Insight Manager, Interoperability, Linux System Administration, Release Management, Reliability Engineering, Cloud Services, Data Logging, Scripting, Enterprise Software Applications, Cloud Platform System, Reliability of Systems, Containerization, Kubernetes, Information Technology, Deployment Automation, Docker, Jenkins - **Published:** June 29, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=98bfc7dab05ad6ce ## About the Role Do you have experience in Tooling?, * Bachelor's degree in Computer Science, Information Technology, Engineering, or related field. * 3+ years of experience in Site Reliability Engineering, DevOps, System Administration, or Production Support. * Experience with Linux/Unix administration and scripting. * Knowledge of monitoring and logging tools. * Experience supporting enterprise applications in cloud environments., * Experience supporting CMS, Medicare, Medicaid, or federal healthcare programs. * Experience with AWS cloud services. * Knowledge of Kubernetes, Docker, and container orchestration. * Familiarity with eSMD, healthcare interoperability, or healthcare data exchange programs. * Experience with CI/CD tools such as Jenkins, GitHub Actions, or Azure DevOps., * Site Reliability Engineering (SRE) * Application Monitoring & Observability * Incident Management & Root Cause Analysis * AWS Cloud Services * Kubernetes & Docker * CI/CD & DevOps Automation * Linux Administration & Scripting * Performance & Capacity Management * Security & Compliance, Must be eligible to obtain and maintain a U.S. Government Public Trust clearance. Candidates must have resided in the United States for at least three (3) of the last five (5) years to satisfy federal background investigation requirements., Candidate must be able to obtain Public Trust clearance and must have lived in the United States for at least three (3) out of the last five (5) years. ## Description The Site Reliability Engineer (SRE) supports the Electronic Submission of Medical Documentation (eSMD) program by ensuring the reliability, availability, performance, and security of applications and infrastructure. The role focuses on system monitoring, automation, incident response, and operational excellence across cloud and on-premises environments supporting CMS healthcare data exchanges., * Monitor application and infrastructure health, performance, and availability. * Implement and maintain observability solutions, dashboards, alerts, and logging. * Support incident management, root cause analysis, and problem resolution. * Automate operational tasks and deployment processes using DevOps tools and scripting. * Collaborate with development, security, and infrastructure teams to improve system reliability and performance. * Support CI/CD pipelines and release management activities. * Manage system capacity planning, scalability, and disaster recovery processes. * Ensure compliance with CMS, HIPAA, FISMA, and federal security requirements. * Support cloud and containerized environments, including Kubernetes and AWS services. * Maintain operational documentation, runbooks, and standard operating procedures. ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Where to Find Entry-Level Software Engineering Jobs](https://www.wearedevelopers.com/magazine/397-where-to-find-entry-level-software-engineering-jobs)