> Markdown version of [/jobs/ext/3101830-principal-site-reliability-engineer-we-have-office-locations-in-cambridge-leeds-and-london](https://www.wearedevelopers.com/jobs/ext/3101830-principal-site-reliability-engineer-we-have-office-locations-in-cambridge-leeds-and-london). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Site Reliability Engineer (we have office locations in Cambridge, Leeds and London) - **Company:** Genomics England - **Location:** Leeds, UK (Remote available) - **Salary:** £107,000.0 - **Contract:** Temporary contract - **Skills:** Application Programming Interfaces (APIs), Amazon Web Services, Application Release Automation, Continuous Integration, Extract Transform Load (ETL), Python (Programming Language), Systems Development Life Cycle, Reliability Engineering, Software Engineering, Web Applications, Datadog, ReactJS, Fast Healthcare Interoperability Resources, Gitlab, Kubernetes, AWS Fargate, Terraform, Api Management, Artifactory, Web Api - **Published:** September 27, 2026 - **Apply:** https://itjobpro.co.uk/job/principal-site-reliability-engineer-we-have-office-locations-in-cambridge-leeds-and-london ## About the Role You are a hands-on engineering leader with deep experience of SRE and platform engineering. You combine strategic thinking with the ability to identify and prioritise the areas where reliability improvements will have the greatest impact. You are a problem-solver who identifies risks, issues, gaps, and dependencies and brings people together to find solutions. You do this through your supportive, empathetic and collaborative behaviours, acting as a coach, mentor, guide or constructive questioner as the situation demands. You are a great communicator, comfortable not just leading your own team, but also engaging across the engineering community and with non-technical stakeholders. Essential Skills and Experience While we recognise the value of relevant qualifications or certifications, we are primarily interested in your real-world experience: Comprehensive knowledge of SRE principles and practices with significant experience of applying these to real-world situations Excellent software engineering skills especially in the context of release automation and other toil-eliminating activities (Python preferred, polyglot ideal) Strong understanding of how architecture and other factors contribute to the overall resilience of systems Extensive experience of platform engineering across CI/CD, Infrastructure as Code, operational monitoring and alerting, backup and recovery etc. Experience in at least one major public cloud (AWS preferred but not essential) Demonstrable ability to lead teams, develop people and coordinate work towards shared outcomes Strong interpersonal skills with a temperament that builds trust and connection within and across squads through open, honest communication Comfortable engaging responsively with teams both remotely and in person when required Ability to navigate rapidly to effective solutions through engaged and inclusive listening, clarity of thought, clear documentation, and succinct presentationDesirable Experience These skills are not essential but if you have either of them, they may prove to be useful: Background in healthcare or bioinformatics Experience in regulated environmentsIf you're an experienced Site Reliability Engineer leader, who thrives on working collaboratively to mature engineering practices, we'd love to hear from you. Join us at Genomics England and make a meaningful impact in the world of genomics. ## Description As Principal Site Reliability Engineer, you will help establish and grow Genomics England's organisation-wide SRE capability. You will be responsible for: Building and leading a small team of Site Reliability Engineers (initially 3 people) Identifying the highest-value opportunities to improve reliability across our platforms Defining standards, patterns and tooling that help engineering teams to build and operate reliable services Developing services and capabilities that reduce operational toil and improve engineering effectiveness Partnering with product and engineering teams to embed SRE principles into their ways of working Influencing engineering strategy, governance and technical direction across the organisation Helping to build a culture of reliability, continuous improvement and operational resilienceThe SRE team will work alongside product squads, helping them adopt reliability practices, improve operations and solve complex technical challenges, rather than operating services on their behalf. In your first year you will establish the initial SRE team, identify the highest-priority reliability challenges across Genomics England, and define the standards, services and practices that will underpin our long-term approach to reliability. The Principal SRE reports to the Director of Engineering within the Technology and Product Directorate. About the Tech Stack The new SRE team will support squads that run a variety of services: most of these are either user-facing web applications (React), backend APIs (Python), bioinformatics pipelines (NextFlow), or data ETL workflows (Prefect, Dremio). These services increasingly run in AWS, though there is still a significant on-premise presence, and they run in a mixture of compute environments, from ECS/Fargate to HPC clusters to (occasionally) Kubernetes. Within the SDLC we have a standard toolchain which includes Terraform for infrastructure-as-code, GitLab for source code and CI/CD, Artifactory for software artefacts, and DataDog for observability. We are working to become interoperable with the wider NHS via open standards like FHIR and GA4GH APIs and increasingly aiming to integrate with their own API Management platform. ## Related Videos - [Watch Tests Go Brrrr! : Getting Started with Cypress in ReactJS](https://www.wearedevelopers.com/videos/282-watch-tests-go-brrrr-getting-started-with-cypress-in-reactjs) - [Web APIs you might not know about](https://www.wearedevelopers.com/videos/281-web-apis-you-might-not-know-about) - [WeAreDevelopers LIVE - Modern DevOps for IoT Devices and More](https://www.wearedevelopers.com/videos/1805-wearedevelopers-live-modern-devops-for-iot-devices-and-more) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [Project Fugu: Extending the web](https://www.wearedevelopers.com/videos/832-project-fugu-extending-the-web) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [Software Engineer Salary in The UK](https://www.wearedevelopers.com/magazine/231-software-engineer-salary-in-the-uk) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)