> Markdown version of [/jobs/ext/2286531-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2286531-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** REVYBE IT RECRUITMENT LIMITED - **Location:** London, UK - **Salary:** £85,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Software as a Service, Cloud Computing, DevOps, Disaster Recovery, Github, Monitoring of Systems, Performance Tuning, Reliability Engineering, Software Engineering, Data Logging, Scripting, Reliability of Systems, Kubernetes, Terraform - **Published:** August 29, 2026 - **Apply:** https://www.totaljobs.com/job/site-reliability-engineer/revybe-it-recruitment-limited-job107912191 ## About the Role * Proven commercial experience working as an SRE, DevOps Engineer, Platform Engineer, or similar * Strong hands-on experience with AWS * Strong experience working with Kubernetes * Excellent experience with Terraform and Infrastructure as Code * Strong experience building and managing GitHub Actions CI/CD pipelines * Solid experience with monitoring and observability tooling * Strong understanding of metrics, logging, tracing, alerting, and system health * Experience troubleshooting complex production environments * Understanding of SLIs, SLOs, SLAs, and error budgets * Experience with incident management and root cause analysis * Good understanding of cloud networking, security, and infrastructure fundamentals * Strong scripting/automation skills * A strong understanding of reliability, scalability, performance, and availability * Excellent communication skills and the ability to work closely with software engineering teams * A proactive mindset and genuine passion for automation and continuous improvement, If you're an experienced SRE, Platform Engineer or DevOps Engineer who enjoys solving complex reliability challenges and wants to have a real impact within a rapidly growing SaaS business, we'd love to hear from you. ## Description This is a hands-on role focused on AWS, Kubernetes, Terraform, observability, monitoring, and automation, working closely with software engineering teams to improve platform reliability and developer experience. You'll have genuine ownership and the opportunity to influence how the platform evolves as the business continues to scale. What You'll Be Doing * Design, build, and maintain highly available and scalable AWS infrastructure * Manage and optimise Kubernetes environments and containerised workloads * Build and maintain infrastructure using Terraform and Infrastructure as Code principles * Develop and optimise CI/CD pipelines using GitHub Actions * Build and improve comprehensive monitoring and observability across the platform * Implement and maintain effective logging, metrics, tracing, alerting, and dashboards * Define and improve SLIs, SLOs, and reliability metrics * Proactively identify and resolve performance, availability, and reliability issues * Lead and contribute to incident response, troubleshooting, and root cause analysis * Automate operational processes and eliminate repetitive manual tasks * Work closely with software engineers to improve deployment processes, system reliability, and developer experience * Help improve platform resilience, scalability, and disaster recovery capabilities * Contribute to capacity planning and performance optimisation as the platform scales * Establish and champion SRE best practices across the wider engineering function ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [JavaScript? No. Java Scripts! - Scripting with Java](https://www.wearedevelopers.com/videos/2094-javascript-no-java-scripts-scripting-with-java) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [The 12 Best Jobs for Software Engineers](https://www.wearedevelopers.com/magazine/401-the-12-best-jobs-for-software-engineers) - [Where to Find Entry-Level Software Engineering Jobs](https://www.wearedevelopers.com/magazine/397-where-to-find-entry-level-software-engineering-jobs)