> Markdown version of [/jobs/ext/1163334-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1163334-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Randstad UK - **Location:** Nottingham, UK (Remote available) - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), ARM Architecture, Bash Shell, Data Infrastructure, Linux, DevOps, Distributed Systems, Elasticsearch, Python (Programming Language), Reliability Engineering, Ansible, Prometheus, Ruby, Scala (Programming Language), Bare Metal, Apache Kafka, Terraform, Data Pipelines, Elk Stack, Golang - **Published:** July 3, 2026 - **Apply:** https://find.jobs/jobs-near-me/apply/ats-redirect/?id=2854467635-2 ## About the Role * 5+ years operating mid-to-large distributed systems on Linux VMs or bare-metal machines. * 2+ years developing in Go, Python, Ruby, Scala, or Bash. * Hands-on experience with Prometheus/Thanos/Cortex, Kafka, the ELK stack, Ansible, or Consul. * Comfortable diving into unfamiliar codebases and participating in an on-call rotation. Keywords: Observability, Monitoring, SRE, Site Reliability Engineering, DevOps, ElasticSearch, ELK, Prometheus, Kafka, Terraform, Linux, Bare Metal ## Description We are looking for a Lead SRE to design, scale, and operate massive-scale observability systems that keep our global services online and performant. You will join an autonomous team of software engineers focused on solving complex data infrastructure challenges., * Scale Prometheus metrics infrastructure to handle 100+ million active series. * Operate large Elasticsearch clusters holding 2000+TB of data. * Grow high-throughput Kafka data pipelines processing hundreds of thousands of events per second. * Build custom alerting workflows and self-service APIs for internal engineering teams. * Provision cloud and private infrastructure using Terraform. ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs)