> Markdown version of [/jobs/ext/3571473-site-reliability-engineering-lead](https://www.wearedevelopers.com/jobs/ext/3571473-site-reliability-engineering-lead). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineering Lead - **Company:** RELX Group plc - **Location:** Washington, DC, United States (Remote available) - **Experience:** Expert - **Salary:** $118,300.0 - $219,800.0 - **Contract:** Permanent contract - **Skills:** User Authentication, Microsoft Azure, Bash Shell, Cloud Computing, Cloud Engineering, Computer Networks, Customer Data Management, DevOps, Domain Name System (DNS), Github, Virtual Private Networks (VPN), Python (Programming Language), Windows PowerShell, Reliability Engineering, Azure Active Directory, Prometheus, TCP/IP, Lexis, Policy as Code, Load Balancing, Cloud Platform System, Autoscaling, Grafana, Kubernetes, Infrastructure Automation Frameworks, Terraform - **Published:** October 3, 2026 - **Apply:** https://dejobs.org/x/x/97E27CE7FCBE4BA4B2634467D39F5BF3/job/ ## About the Role * Expert knowledge of Kubernetes, including cluster architecture, upgrades, autoscaling, security hardening, and troubleshooting at scale * Expert experience with Terraform, including modular IaC design, state management, multi-environment provisioning, and policy-as-code * Deep knowledge of Azure Cloud, including compute, networking, identity (AAD), storage, and cost optimization * Experience designing and scaling CI/CD pipelines using GitHub Actions, release strategies, and rollback automation * Experience with observability platforms including Prometheus, Grafana, OpenTelemetry, and SLO/SLA/error-budget management * Strong automation skills, focused on eliminating toil through self-healing systems and infrastructure automation * Advanced proficiency in Python, Bash, and/or PowerShell for tooling and automation * Deep understanding of networking concepts including TCP/IP, DNS, load balancing, VPNs, and cloud-native networking * Experience in SRE, DevOps, or Infrastructure roles, including experience leading engineering teams * Proven track record leading incident response and driving reliability improvements ## Description Are you passionate about building reliable, scalable platforms and helping engineering teams thrive? Do you enjoy leading high-performing teams while driving automation, resilience, and operational excellence across modern cloud environments? About the Business LexisNexis Risk Solutions is the essential partner in the assessment of risk. Within our Business Services vertical, we offer a multitude of solutions focused on helping businesses of all sizes drive higher revenue growth, maximize operational efficiencies, and improve customer experience. Our solutions help our customers solve difficult problems in the areas of Anti-Money Laundering/Counter Terrorist Financing, Identity Authentication & Verification, Fraud and Credit Risk mitigation and Customer Data Management. You can learn more about LexisNexis Risk at https://risk.lexisnexis.com/ About Our Team You will be joining the Core SRE Team in Business Services, a team that oversees all the applications and infrastructure in the biggest business unit in LexisNexis Risk Solutions. We build cloud environments, migrate on prem applications to the cloud, work with self-hosted and 3rd party solutions. The successful candidate is a self-starter who assesses the situation, collaboratively develops a solution, and takes the initiative to improve performance, cost, and reliability at each opportunity. About the Role This is a professional management level role. Individuals are required to provide line management to a small to medium sized team of engineers, including authority over performance management, pay, and recruitment. They will ensure that tasks and projects are prioritized appropriately and provide support to team members when tasks are blocked. They will support engineers in their personal development and ensure they are working within the SRE framework. They will lead the post-mortem reviews and the timely production of RCAs. They address issues with impact beyond their own team based on knowledge of related disciplines. Responsibilities * Manage, mentor, and grow a team of SREs; conduct 1:1s, performance reviews, and career development planning * Own hiring, onboarding, and team capacity/resourcing decisions * Set team goals, prioritize backlog, and drive planning * Foster a blameless post-incident culture and cross-team collaboration with Dev, Security, and Product * Lead reliability initiatives across infrastructure and services * Drive incident response activities and continuous service improvement * Champion automation and operational excellence across the platform * Support the development of scalable, secure, and resilient cloud-native environments ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Turning Container security up to 11 with Capabilities](https://www.wearedevelopers.com/videos/718-turning-container-security-up-to-11-with-capabilities) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)