SRE Technical Lead

1 Hr Ago By 83zero Ltd
Wokingham, United Kingdom
2 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Senior
Compensation
£ 100K

Job location

Remote
Wokingham, United Kingdom

Tech stack

Cloud Computing
Databases
PostgreSQL
Openshift
Red Hat Enterprise Linux - RHEL
Reliability Engineering
Site Reliability Engineering Practices
Prometheus
Service Design
Datadog
Istio
Grafana
Multi-Cloud
Kubernetes

Job description

This is a senior technical leadership role where you'll act as the technical authority for Site Reliability Engineering, working closely with stakeholders, engineering teams and delivery partners to drive reliability across large-scale cloud platforms. Essential Security Requirements

To be considered for this role, you must:

  • Hold active SC (Security Check) clearance.
  • Be a sole UK national.

Unfortunately, candidates who do not meet both of these essential requirements cannot be considered., As the SRE Technical Lead, you will:

  • Define and drive the SRE strategy, standards, SLAs, SLOs and error budgets.
  • Embed reliability engineering principles into platform and service design.
  • Lead the adoption of core SRE practices including reliability reviews, operational readiness and toil reduction.
  • Drive automation across monitoring, incident response, recovery and remediation.
  • Govern reliability-focused Infrastructure as Code, CI/CD pipelines and operational tooling.
  • Identify and remove systemic causes of operational overhead while improving scalability, resilience and operability.
  • Act as the senior technical escalation point for major incidents and high-risk releases.
  • Lead blameless post-incident reviews and ensure measurable service improvements.
  • Define and oversee observability, monitoring and capacity management practices.
  • Ensure SRE approaches align with security, governance and compliance requirements.
  • Mentor and coach senior engineers, helping to improve SRE maturity across engineering teams.

This role provides technical leadership through influence rather than direct line management and works closely with Cloud, Platform, Security and Operations teams.

Requirements

We're looking for an experienced SRE Technical Lead to take ownership of the reliability, availability and operational excellence of critical platforms within complex, multi-vendor environments., You'll have strong technical expertise gained within enterprise-scale environments, including:

  • Deep knowledge of Kubernetes and OpenShift.
  • Experience designing and supporting hybrid and multi-cloud platforms.
  • Experience with service mesh technologies such as Istio.
  • Strong hands-on experience with observability tooling including Prometheus, Grafana, Loki, Tempo and OpenTelemetry.
  • Infrastructure as Code and GitOps expertise using tools such as Helm, Kustomize, ArgoCD and Tekton.
  • Experience building and improving CI/CD pipelines with a focus on reliability engineering.
  • Familiarity with Red Hat ACM/ACS, Submariner networking and enterprise databases such as PostgreSQL.
  • A proven track record of providing technical leadership across complex, multi-vendor environments.

Benefits & conditions

  • Salary up to £100,000
  • 5% annual bonus
  • Hybrid working model
  • Opportunity to lead reliability engineering across large-scale, business-critical platforms
  • Exposure to modern cloud-native technologies and complex enterprise environments

Apply for this position