> Markdown version of [/jobs/ext/3042848-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3042848-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Biometric Talent - **Location:** Manchester, UK - **Salary:** £40,000.0 - £65,000.0 - **Contract:** Permanent contract - **Skills:** JavaScript (Programming Language), Application Programming Interfaces (APIs), Artificial Intelligence, Business Analytics Applications, Information Technology Operations, Python (Programming Language), Reliability Engineering, Ansible, Shell Script, Software Engineering, Large Language Models, Grafana, Reliability of Systems, Kubernetes, Performance Monitor, Software Coding, Terraform, Splunk, New Relic (SaaS), Pagerduty, Golang - **Published:** September 24, 2026 - **Apply:** https://itjobpro.co.uk/job/site-reliability-engineer-6 ## About the Role Our client is looking for a strong software engineer with SRE experience, or someone who has a strong software engineering background and has moved into Site Reliability Engineering. You'll ideally have: * Strong software engineering experience, particularly with Python or Golang * Experience with monitoring, alerting and observability * Knowledge of OpenTelemetry and modern observability practices * Experience establishing proactive monitoring and alerting for complex platforms * Strong understanding of SRE principles, including SLIs and SLOs * Experience with modern software development practices and lifecycles * Proficiency in shell scripting * Experience with Infrastructure as Code, automation and orchestration, ideally using Terraform and Ansible * Experience with tools such as Grafana, Splunk, New Relic and PagerDuty * Experience working within large-scale, 24/7 enterprise environments where availability and stability are critical * Strong incident management, troubleshooting and root cause analysis experience They're also looking for an AI-native approach to engineering, with hands-on experience using LLM platforms and coding assistants to improve productivity and quality. Experience or interest in using AI for telemetry, predictive insights and root-cause analysis would be particularly relevant. ## Description We're supporting our client with the appointment of an experienced Site Reliability Engineer (SRE) to join their Platform Engineering team. This is a software-focused SRE position, where your software engineering skills will be used to solve operational problems and improve the reliability, observability and performance of a large-scale production environment. Working closely with Development, Platform Delivery and IT Operations teams, you'll help build the tooling, automation, monitoring and practices that keep critical systems reliable and resilient. If you're a software engineer who enjoys solving complex operational problems, writing code and using engineering to improve system reliability, this could be an interesting next step. How you'll spend your day You'll work across software engineering, observability, automation and incident management, with a focus on proactively improving the health and performance of critical services. Key responsibilities include: * Writing and contributing to code that improves service reliability and observability * Developing tools, operational APIs and automation to improve system management * Establishing proactive monitoring and alerting across complex platforms * Implementing service instrumentation using OpenTelemetry * Building sophisticated dashboards using Grafana, Splunk and New Relic * Automating manual processes and reducing operational toil * Working with Infrastructure as Code and orchestration technologies * Supporting live incident resolution and contributing to post-mortem analysis * Carrying out root cause analysis and implementing effective remediation * Driving initiatives to improve system reliability, performance and observability * Maintaining and administering existing monitoring and analytics platforms * Working with IT Operations to provide critical tooling and capabilities * Sharing knowledge and mentoring colleagues on new technologies and practices The role also has a strong focus on AI-enabled engineering. You'll use AI tools, LLM platforms and coding assistants in your day-to-day work to improve productivity, reduce toil and explore new approaches to autonomous operations, telemetry and system health. Technology environment The successful candidate will work across a modern engineering environment, with technologies including: Python, Golang and JavaScript OpenTelemetry Grafana Splunk New Relic PagerDuty Ansible Terraform Infrastructure as Code Shell scripting Automation and orchestration platforms AI tools, LLM platforms and coding assistants, Should we both wish to proceed, we will submit your details to the client and be in touch regarding the outcome and any further steps. ## Related Videos - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Scoring 2000 Products per Request: Performance Pitfalls in Golang](https://www.wearedevelopers.com/videos/2073-scoring-2000-products-per-request-performance-pitfalls-in-golang) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london)