> Markdown version of [/jobs/ext/22714-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/22714-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** The Reward - **Location:** UK - **Contract:** Permanent contract - **Skills:** PHP (Programming Language), Amazon Web Services, Bash Shell, Continuous Integration, DevOps, Python (Programming Language), Reliability Engineering, Site Reliability Engineering Practices, SQL Databases, Systems Integration, Datadog, Kubernetes, Atlassian Tools, Terraform - **Published:** May 15, 2026 - **Apply:** https://www.totaljobs.com/job/site-reliability-engineer/reward-gateway-job107309035 ## About the Role We have a highly talented team who live up to our shared values, bringing to life "Entrepreneurial Spirit". We love to "Push the Boundaries", but importantly show "Respect" in how we go about our work. Our "Speak Up" and "Be Human" culture is at the heart of what it is to work with us, and we encourage everyone to bring innovation and 'Imagination' to the work they do., * Proven experience in DevOps or SRE, with a keen interest in growing as a Site Reliability Engineer * Experience with AWS or other cloud providers * Enterprise experience in HA environments * Automation skills through Terraform, Python, Bash or similar * Wide-reaching SRE skills and a deep understanding of SRE practices * A strong understanding of SQL, PHP, Kubernetes, CI/CD * Observability product experience (e.g., Datadog) * Managing services using SLI/SLO & Error Budgets * Ability to work both independently and as part of a team * Ability to work under pressure and be highly reliable * Adaptability and flexibility to change in a fast-moving environment * An ability to learn new tools and processes quickly and impart that knowledge ## Description Reward Gateway, together with Edenred, are a global market leader in benefits and employee engagement. We help our clients and their leaders to transform employee experience that will attract, engage and retain top talent through employee benefits, strategic reward and recognition, well-being, and much more. With our shared missions of 'Making the World a Better Place to Work" and 'Enriching connections, For good'. You'll be contributing to improving employee engagement and building better, stronger and more resilient organisations to improve people's daily lives. Our shared mission guides our every action and charts a sustainable path to a better future., Collaboration, connection as a team, and strong internal relationships are part of the "RG Magic" that makes our culture thrive. Our teams work from our Dean Street office two days per week. What You'll be Doing: * Integrating tightly with our Product Engineering teams * Following SRE practices and maintaining high standards of compliance * Implementing a new standard of observability utilising SLI/SLO/Error Budgets * Continually evolving our observability platforms for greater coverage * Using a code-first approach to build and changes to reduce TOIL * Advocating a strong focus on availability, reliability and uptime * Liaising and embedding with the Engineering teams for the constant evolution of metrics * Working towards planned roadmap goals * Actively taking part in the daily stand-ups and keeping sprints on track * Keeping up-to-date documentation in the JIRA & Confluence tools * Taking part in SRE Incident Management processes * Acting as a key Incident Commander within the Incident Management process * Taking part in SRE On Call * Ensuring a focus on cost efficiency for the platforms & services * Working with team members to foster collaboration and ongoing communication with stakeholders ## Related Videos - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [The OpenTelemetry mistakes I keep seeing (and how to stop making them)](https://www.wearedevelopers.com/videos/100158-the-opentelemetry-mistakes-i-keep-seeing-and-how-to-stop-making-them) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Implementing Feature Environments with AWS and Terraform](https://www.wearedevelopers.com/videos/531-implementing-feature-environments-with-aws-and-terraform) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Fullstack Developer Salary UK](https://www.wearedevelopers.com/magazine/251-fullstack-developer-salary-uk) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk)