SRE

SR2
London, UK
6 days ago
Apply on www.careerboard.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Compensation
£156,000.0 - £195,000.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Microsoft Azure Bash Shell Cloud Computing Cloud Engineering Databases Continuous Integration Linux Github Python (Programming Language) PostgreSQL OpenShift
+17 more
Reliability Engineering Prometheus Zero Trust Network Access Software Engineering Datadog Google Cloud Cloud Platform System Delivery Pipeline Grafana Git Flow Kubernetes Infrastructure Automation Frameworks Apache Kafka Terraform Devsecops Docker Golang

Job description

SR2 is supporting a major long term programme and looking for an experienced Site Reliability Engineer (SRE) to join the Production Engineering team. This function underpins the reliability, security, and performance of all live environments, from production systems to critical customer deployments. You’ll apply a software engineering mindset to operational problems, building automation, scalability, and resilience into cloud-native infrastructure. Beyond supporting live systems, this team also acts as a centre of excellence, guiding project teams in adopting best practices across DevSecOps, observability, and cost optimisation., * Build, maintain, and support production and demo environments

  • Automate infrastructure provisioning and deployment workflows (Terraform, GitHub Actions, GitOps)
  • Package and deploy applications to customer environments
  • Implement and optimise observability tooling (Prometheus, Grafana, Loki)
  • Support incident response, monitoring, and backup/recovery planning
  • Mentor project teams in DevSecOps practices and environment management
  • Ensure cloud environments are secure, performant, and cost-optimised

Requirements

  • Cloud Engineering: AWS/Azure/GCP, Linux, Terraform (IaC)
  • Containers: Kubernetes, Docker, Helm (OpenShift a plus)
  • Observability: Prometheus, Grafana, Loki (network visualisation desirable)
  • CI/CD & GitOps: GitHub Actions, ArgoCD/Flux
  • Security: Cloud access models, Zero Trust principles
  • Programming: Python or Golang preferred; Bash Scripting
  • Databases & Pipelines: PostgreSQL optimisation, Kafka integration
  • Excellent problem-solving skills, with the ability to troubleshoot complex issues quickly

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerboard.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

1:29 min

Recommended community resources for cloud engineers

Piet Van Dongen · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

Videos

See all

Related articles

See all