> Markdown version of [/jobs/ext/1992298-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1992298-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** LexisNexis Risk Solutions Group - **Location:** Boca Raton, FL, United States (Remote available) - **Experience:** Expert - **Salary:** $104,900.0 - $174,700.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Confluence, JIRA, Microsoft Azure, Linux, DevOps, Github, Reliability Engineering, Grafana, Git Flow, Kubernetes, Terraform, Docker, Servicenow - **Published:** August 8, 2026 - **Apply:** https://www.careerbuilder.com/job-details/senior-site-reliability-engineer-ii-boca-raton-fl--c6c86b5b-b2ca-4862-89dc-4d41a934fc48 ## About the Role * 5+ years of hands-on experience in SRE, DevOps, or Infrastructure Engineering roles * Strong production experience in AWS * Required: Significant hands-on experience with Terraform in real-world environments * Experience operating monitoring and uptime platforms such as Grafana, Pingdom, and Uptrends * Strong Linux systems, networking, and troubleshooting skills * Experience supporting production systems through incident response and on-call rotations * Proficiency with GitHub and modern Git workflows * Experience building or maintaining CI/CD pipelines with Azure DevOps * Familiarity with ITSM and incident workflows using ServiceNow * Strong written communication skills with experience documenting systems and processes in Confluence * Ability to work independently in a remote or hybrid environment Preferred Qualifications * Experience defining and operating against SLOs and error budgets * Infrastructure-as-Code best practices beyond Terraform (modules, testing, CI integration) * Experience with containers and orchestration (Docker, Kubernetes) * Experience supporting large-scale, high-availability production systems * Prior experience mentoring engineers or serving as a technical lead ## Description We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory role you will be directly involved in designing infrastructure, writing Terraform, improving observability, and responding to real production incidents., * Design, build, and operate highly available, scalable systems in AWS * Write, maintain, and review Terraform to provision and manage infrastructure * Own and improve monitoring, alerting, and observability using Grafana, Pingdom, and Uptrends * Participate in a rotating on-call schedule, responding to production incidents and driving issues to resolution * Lead incident response, root cause analysis, and post-incident reviews with a focus on prevention and automation * Define and manage SLOs, SLIs, and error budgets * Build and improve CI/CD pipelines and operational workflows using Azure DevOps and GitHub * Work directly with application teams to improve reliability, performance, and deployability * Automate manual operational tasks to reduce toil * Maintain clear, actionable runbooks and documentation in Confluence * Track work, incidents, and operational improvements using Jira and ServiceNow * Mentor other engineers and help set SRE standards and best practices ## Related Videos - [Collaboration Quantified: Lessons from Open Source Developer Networks](https://www.wearedevelopers.com/videos/1422-collaboration-quantified-lessons-from-open-source-developer-networks) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Applying Agile Principles to Incident Management ](https://www.wearedevelopers.com/videos/101-applying-agile-principles-to-incident-management) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [What is Software Engineering?](https://www.wearedevelopers.com/magazine/289-what-is-software-engineering)