> Markdown version of [/jobs/ext/3330359-senior-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3330359-senior-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer - **Company:** Spectrum IT Recruitment - **Location:** Southampton, UK - **Experience:** Expert - **Salary:** £60,000.0 - £70,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Amazon Web Services, Amazon Elastic Compute Cloud, Systems Engineering, Bash Shell, C Sharp (Programming Language), Cloud Computing, Configuration Management, Continuous Integration, DevOps, Distributed Systems, Amazon DynamoDB, Monitoring of Systems, Python (Programming Language), Windows PowerShell, Reliability Engineering, Ansible, Prometheus, Datadog, Circleci, Scripting, Google Cloud, Enterprise Software Applications, System Availability, Grafana, Cloudformation, Containerization, Gitlab-ci, Kubernetes, Infrastructure Automation Frameworks, Cloudwatch, Puppet, Rundeck, Terraform, Splunk, Docker, Pagerduty, Elk Stack, Jenkins, Golang, Microservices - **Published:** September 9, 2026 - **Apply:** https://www.adzuna.co.uk/jobs/details/5876126152 ## About the Role * Practical experience managing large-scale Kubernetes clusters; certifications in Kubernetes are a strong bonus * Hands-on familiarity with the Grafana Observability Suite, including tools like Loki, Mimir, and Tempo * Background in administering or developing with popular monitoring and automation tools such as Splunk, Datadog, PagerDuty, or Rundeck * Experience using configuration management platforms like Ansible, Puppet, or Chef * Professional certifications in cloud DevOps, such as AWS Certified DevOps Engineer or Google Cloud Professional DevOps Engineer, or similar credentials Do You Have What It Takes? * 3-6 years of hands-on experience in a similar role, with a strong emphasis on systems engineering, automation, and service reliability * Proficient in at least one programming language such as Python, Go, Java, or C#, along with scripting skills in Bash or PowerShell * Solid grasp of cloud platforms like AWS, including an understanding of how core services like EC2, ECS, Lambda, and DynamoDB operate under reliability constraints * Practical experience using infrastructure-as-code tools like CloudFormation or Terraform * In-depth knowledge of CI/CD principles and hands-on experience with tools such as Jenkins, GitLab CI/CD, or CircleCI * Strong understanding of containerization (e.g., Docker, Kubernetes) and microservices architecture * Skilled in using observability and monitoring tools such as Prometheus, Grafana, ELK stack, or AWS CloudWatch * Excellent analytical and troubleshooting abilities, especially within complex distributed systems * Proven experience handling incident management and conducting blameless postmortems, including leading cross-functional teams through resolution and communication during critical outages ## Description The company deliver cutting-edge enterprise software solutions across both cloud and on-premises environments, empowering organisations to enhance customer experiences, maintain regulatory compliance, and proactively fight fraud. The company are trusted by businesses worldwide to drive seamless, intelligent customer interactions. In this role, you'll oversee the production environment by ensuring system availability and maintaining a comprehensive perspective on overall health. You'll develop tools and software to support and streamline the management of platform infrastructure and key applications. A major focus will be enhancing the dependability, performance, and delivery speed of our software products. You'll also be responsible for analysing and fine-tuning system performance to anticipate user demands and drive innovation. Additionally, you'll take the lead in providing operational support and technical oversight for several large-scale distributed applications. How You'll Contribute: * Monitor and interpret system and application metrics to fine-tune performance and troubleshoot issues effectively * Collaborate closely with developers to enhance service quality through thorough testing and structured release practices * Engage in architectural discussions, manage platform operations, and contribute to capacity forecasting * Design and implement automated solutions to build resilient, scalable systems * Maintain a strong focus on delivering new features while ensuring stability and adherence to service level goals ## Related Videos - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [7 Most Popular Web Developer Jobs in Europe](https://www.wearedevelopers.com/magazine/163-7-most-popular-web-developer-jobs-in-europe) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers)