Senior Site Reliability Engineer II ** Remote EST Preferred

RELX Group plc
Horsham, PA, United States
about 1 month ago
Apply on relx.wd3.myworkdayjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours

Tech stack

Artificial Intelligence Microsoft Azure Continuous Integration Disaster Recovery Github Log Analysis Performance Tuning Systems Development Life Cycle Reliability Engineering Azure DevOps Pipelines Datadog Scripting
+7 more
Cloud Platform System GitHub Copilot Infrastructure Automation Frameworks Deployment Automation Azure AKS Software Version Control Serverless Computing

Job description

We’re looking for a Senior Site Reliability Engineer II to serve as a go-to technical resource for our Life Sciences SRE team, part of a globally distributed SRE organization. This is a developed professional role for an experienced SRE who uses knowledge of internal and external issues to improve services and processes, owns complex reliability and toil-reduction projects, and acts as an escalation point during major incidents. You’ll help drive day-to-day operations for the Life Sciences SRE team, working within priorities set by the SRE Manager, and partner closely with the Senior Software Architect for Life Sciences to modernize processes, tooling, and development strategy. You’ll also work across the broader Reed Tech organization with SREs from other business units to converge on common tools, standards, and ways of working. As part of our broader technology strategy, the team is actively investing in AI-driven development and operations - candidates with experience or strong interest in applying AI tooling to engineering workflows are especially encouraged to apply.

Requirements

  • Help drive day-to-day operations for the Life Sciences SRE team, supporting priorities set by the SRE Manager
  • Partner with SREs and contractors across the Life Sciences SRE team to deliver reliability and toil-reduction initiatives
  • Partner closely with the Senior Software Architect for Life Sciences to modernize processes, tooling, and development strategy
  • Collaborate with SREs across the broader Reed Tech organization to converge on common tools, standards, and practices
  • Influence the setting of service level objectives (SLOs) across Life Sciences production systems
  • Participate in resolving major incidents as an escalation point and guide restoration efforts
  • Troubleshoot and resolve complex systems and application issues across development and production environments, partnering with development teams as needed
  • Improve the SRE framework and contribute to shared SRE knowledge documentation
  • Create disaster recovery plans
  • Mentor and prepare junior SREs for on-call readiness
  • Explore and champion AI-assisted development and operational tooling as part of the team’s broader AI strategy
  • Cloud Platform: Microsoft Azure
  • Compute: VMs/VMSS, App Service, Azure Functions, with growing use of Azure Kubernetes Service (AKS)
  • CI/CD: Azure DevOps Pipelines, with source control transitioning to GitHub (GitHub Actions under evaluation)
  • Observability: Azure Monitor and Log Analytics (primary), Datadog (supplemental), * Supports platform modernization, CI/CD and SDLC improvements, adoption of standard platform capabilities, and operational health reporting.
  • Technical Influence & Collaboration
  • Acts as a trusted technical advisor, contributes to modernization initiatives, drives alignment on standards and tooling, and mentors team members.
  • Reliability Culture
  • Promotes SRE best practices, coaches engineers, identifies skill gaps, and champions automation and toil-reduction initiatives.

Work in a Way That Works for You We promote a healthy work/life balance across the organisation. We offer an appealing working prospect for our people. With numerous wellbeing initiatives, shared parental leave, study assistance and sabbaticals, we will help you meet your immediate responsibilities and your long-term goals.

Requirements

  • Experience using AI-assisted coding tools (e.g., Claude, GitHub Copilot, or Codex) in day-to-day development or scripting work
  • Interest or hands-on experience applying AI/ML to observability, automation, or incident response (e.g., anomaly detection, automated triage, AI-assisted RCA drafting)
  • Comfortable experimenting with and evaluating emerging AI tooling as part of continuous improvement

Responsibilities

  • Observability
  • Strong understanding of full-stack observability, incident analysis, and performance optimization. Able to implement reusable dashboards, observability standards, SLOs, and error budgets.
  • Incident Management
  • Experienced in on-call readiness, mentoring, major incident response, blameless post-mortems, and reducing alert fatigue through root-cause analysis.
  • Design for Reliability
  • Advanced knowledge of high-availability systems, resiliency patterns, deployment strategies, and recovery practices. Provides guidance on reliability improvements and SRE standards.
  • Disaster Recovery
  • Experienced with failover testing, production recovery, and automating recovery processes using Infrastructure-as-Code and configuration management tools.

Benefits & conditions

We know that your wellbeing and happiness are key to a long and successful career. These are some of the benefits we are delighted to offer:

  • Health Benefits: Comprehensive, multi-carrier program for medical, dental and vision benefits
  • Retirement Benefits: 401(k) with match and an Employee Share Purchase Plan
  • Wellbeing: Wellness platform with incentives, Headspace app subscription, Employee Assistance and Time-off Programs
  • Short-and-Long Term Disability, Life and Accidental Death Insurance, Critical Illness, and Hospital Indemnity
  • Family Benefits, including bonding and family care leaves, adoption and surrogacy benefits
  • Health Savings, Health Care, Dependent Care and Commuter Spending Accounts
  • Up to two days of paid leave each to participate in Employee Resource Groups and to volunteer with your charity of choice

About the company

For over 50 years, LexisNexis Reed Technology has partnered with the U.S. Patent and Trademark Office (USPTO) to deliver secure, scalable, and high-quality patent data processing solutions. Our work transforms complex, unstructured patent submissions into standardized, searchable outputs that power examiner workflows and public dissemination.

We operate in a highly regulated environment with strict security requirements, large-scale data volumes, and complex business rules. Our focus is on modernizing legacy workflows through automation, AI, and platform transformation to improve efficiency, accuracy, and cost effectiveness., LexisNexis Legal & Professional® provides legal, regulatory, and business information and analytics that help customers increase their productivity, improve decision-making, achieve better outcomes, and advance the rule of law around the world. As a digital pioneer, the company was the first to bring legal and business information online with its Lexis® and Nexis® services.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on relx.wd3.myworkdayjobs.com
Prepare application

Good distractions

Loading talks and stories from around this role…