> Markdown version of [/jobs/ext/2813052-site-reliability-engineer-sre-financial-wellbeing](https://www.wearedevelopers.com/jobs/ext/2813052-site-reliability-engineer-sre-financial-wellbeing). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer (SRE) - Financial Wellbeing - **Company:** Lloyds Banking Group - **Location:** London, UK - **Experience:** Expert - **Salary:** £70,929.0 - £78,810.0 - **Contract:** Permanent contract - **Skills:** Clean Code Principles, Agile Methodology, JIRA, Microsoft Azure, Cloud Computing, Github, Key Management, Reliability Engineering, Workflow Management Systems, Scripting, File Transfer Protocol (FTP), Cloud Platform System, Containerization, Terraform, Docker, Jenkins - **Published:** September 10, 2026 - **Apply:** https://www.collegerecruiter.com/job/2859573270-site-reliability-engineer-sre-financial-wellbeing ## About the Role * Proven experience as a Site Reliability Engineer in cloud environments (GCP or AWS). * Understanding of SRE principles including SLIs, SLOs, error budgets, and toil reduction. * Strong scripting and infrastructure-as-code (IaaC) skills (Terraform, Harness, GitHub). * Demonstrable experience in the Agile ways of working that focuses on delivering customer value and applying the Agile mindset; familiarity with tools like Jira. * Ability to lead incident response and drive service improvements. * Strong collaboration and mentoring skills., * Azure cloud environment experience, including connectivity, data buckets, secrets management, migration, and governance challenges. * Familiarity with containerisation and orchestration tools like Docker, Jenkins, GitHub, and Terraform. * Secure programming practices and experience of secure file transfer protocols, risk remediation, and audit actions. * Technical operations and service engineering. ## Description The Recoveries & Application Management Lab are building new systems and solutions. The role offers the chance for someone to engage and share their experience, serving as a mentor in the team's decision-making processes related to site reliability that are effective and efficient. Once the SRE support is evaluated, the selected person will join the Platform Engineering team to establish the SRE role by managing the cloud environment, developing solution designs, explaining incident management practices, conducting root cause analysis, and handling business changes. The successful candidate will be required to balance the demands of the LBG business and both stability and scalability requirements to reach a successful and scalable outcome. Day to Day * Create documentation that details the establishment of the SRE function within the platform, supported by procedures that outline the guidelines to be followed through the incorporation of existing documentation. * Provide a framework in which to operate the cloud systems. * Lead the transition to cloud infrastructure and improve observability across systems. * Identify and eliminate toil through automation. * Manage incidents and post-mortems to improve service reliability. * Mentor engineers and support team development. * Collaborate with Product Owners to balance operational and development priorities. ## Related Videos - [Improving quality with Agentic AI with Rovo Dev and Xray](https://www.wearedevelopers.com/videos/2005-improving-quality-with-agentic-ai-with-rovo-dev-and-xray) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Applying Agile Principles to Incident Management ](https://www.wearedevelopers.com/videos/101-applying-agile-principles-to-incident-management) - [Collaboration Quantified: Lessons from Open Source Developer Networks](https://www.wearedevelopers.com/videos/1422-collaboration-quantified-lessons-from-open-source-developer-networks) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Best Companies to work for in London: Top 25 Companies in 2023](https://www.wearedevelopers.com/magazine/187-best-companies-to-work-for-in-london-top-25-companies-in-2023) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk)