> Markdown version of [/jobs/ext/2813307-senior-site-reliability-engineer-application-api-focused](https://www.wearedevelopers.com/jobs/ext/2813307-senior-site-reliability-engineer-application-api-focused). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer (Application / API Focused) - **Company:** Xpertise Recruitment - **Location:** Greater London, UK - **Experience:** Expert - **Salary:** £100,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Amazon Web Services, Application Layers, Cloud Engineering, Continuous Integration, Distributed Systems, Reliability Engineering, Site Reliability Engineering Practices, Prometheus, Datadog, Backend, Kubernetes, Splunk, Microservices - **Published:** September 10, 2026 - **Apply:** https://www.collegerecruiter.com/job/2840587008-senior-site-reliability-engineer-application--api-focused ## About the Role * Proven experience operating as an SRE within digital product environments * Strong understanding of API architectures, microservices and distributed systems behaviour * Hands-on experience defining and implementing SLIs, SLOs and error budgets * Deep observability exposure (e.g. Datadog, Splunk, Prometheus, tracing/APM platforms) * Experience working closely with application engineering teams, not just infrastructure teams * Background in high-availability, customer-facing systems where outages have commercial impact * Cloud-native exposure (AWS preferred) with practical understanding of Kubernetes environments, This role is best suited to engineers who care deeply about production behaviour, customer experience in failure scenarios, and reliability as a first-class product feature, rather than engineers focused purely on infrastructure provisioning or CI/CD enablement. ## Description We are hiring a Senior SRE to support a large-scale digital organisation undergoing a major commercial re-platforming across web and mobile channels. This role sits much closer to the application layer than traditional infrastructure SRE positions. You will work directly with product and engineering teams across customer-facing platforms (web, mobile, payment journeys, APIs) to improve reliability, resilience, and service behaviour in production. This is not a ticket-driven operational role and not a pure platform engineering post. It is about embedding measurable reliability into distributed systems at service level. What You'll Be Doing * Embed SRE practices across API and microservices-based architectures * Define and own meaningful SLIs/SLOs aligned to customer journeys and business-critical flows * Improve service reliability through proactive observability, tracing, telemetry and alert tuning * Partner closely with backend and platform engineers to reduce systemic failure modes * Lead and contribute to incident response, post-incident reviews and resilience improvements * Move the organisation from symptom-based alerting to customer-impact driven diagnostics * Contribute to release safety, progressive deployments and production guardrails ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Debugging in the Dark](https://www.wearedevelopers.com/videos/1658-debugging-in-the-dark) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [Nest.js - TypeScript in the backend can also be clean](https://www.wearedevelopers.com/videos/1033-nest-js-typescript-in-the-backend-can-also-be-clean) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [The 12 Best Jobs for Software Engineers](https://www.wearedevelopers.com/magazine/401-the-12-best-jobs-for-software-engineers)