> Markdown version of [/jobs/ext/1286901-senior-site-reliability-engineer-sre](https://www.wearedevelopers.com/jobs/ext/1286901-senior-site-reliability-engineer-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer (SRE) - **Company:** Tes - **Location:** Sheffield, UK - **Experience:** Expert - **Salary:** £90,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, JIRA, Microsoft Azure, Bash Shell, Cloud Computing, Cloud Engineering, Configuration Management, Continuous Integration, Data Security, Software Design Patterns, DevOps, Monitoring of Systems, Issue Tracking Systems, Python (Programming Language), Uptime, Reliability Engineering, Site Reliability Engineering Practices, Ansible, Prometheus, Software Engineering, Data Logging, Scripting, Google Cloud, Cloud Platform System, Travis Ci, Grafana, Infrastructure as Code (IaC), Containerization, Gitlab-ci, Kubernetes, Deployment Automation, Terraform, Docker, Pagerduty, Jenkins, Microservices - **Published:** July 16, 2026 - **Apply:** https://find.jobs/jobs-near-me/apply/ats-redirect/?id=2876098380-2 ## About the Role * Proven experience in a SRE/DevOps/Platform role, with a strong background in both software development or operations. * Knowledge of CI/CD tools (e.g., Jenkins, GitLab CI/CD, Travis CI). * Proficiency in scripting and automation (e.g., Bash, Python, Ansible). * Strong experience with containerization and orchestration technologies (e.g., Docker, Kubernetes). * Strong hands-on experience of at least one major public cloud platforms (e.g., AWS, Azure, GCP). * Strong problem-solving and troubleshooting abilities in a timebound situation (Major incidents). * Clear communication and incident management experience. * Demonstrable strong hands-on experience with Terraform. * Knowledge of microservices architecture. * Familiarity with security best practices and tools. * Demonstrable experience of monitoring / observability tools preferred Grafana, Prometheus, PagerDuty, uptime. Knowledge * Cloud Platforms: Strong knowledge of AWS, Azure, or GCP, including cloud architecture, services, and security models. * Containerization & Orchestration: In-depth understanding of Docker and Kubernetes for deploying and managing containerized applications. * Infrastructure as Code (IaC): Knowledge of IaC frameworks, particularly Terraform, to manage cloud infrastructure via code. * Microservices Architecture: Familiarity with microservices design patterns and deployment strategies in a cloud-native environment. * Monitoring & Observability: Understanding of monitoring, logging, and alerting tools (e.g., Prometheus, Grafana, ELK) to ensure system performance and issue tracking. Skills * CI/CD Tools: Hands-on experience with Jenkins, GitLab CI/CD, Travis CI, or similar tools for building CI/CD pipelines. * Scripting & Automation: Proficiency in scripting languages like Bash and Python, along with automation tools such as Ansible for managing configurations and deployments. * Containerization & Orchestration: Practical skills in deploying and managing containers using Docker and orchestrating workloads using Kubernetes. * Cloud Platform Management: Expertise in managing and scaling cloud environments on AWS, Azure, or GCP, leveraging services for compute, storage, networking, and security. * Infrastructure as Code (IaC): Skilled in using Terraform to automate provisioning and management of cloud infrastructure. * Troubleshooting & Problem Solving: Strong analytical skills for identifying and resolving complex system issues, especially in production environments. * Collaboration & Communication: Excellent ability to work under pressure e.g. in a Major incident., * Certifications (Preferred): Holding certifications such as AWS Certified DevOps Engineer, CKA (Certified Kubernetes Administrator), or other relevant credentials. ## Description As a Senior SRE Engineer, you will be pivotal in designing and implementing best SRE practices while fostering a culture of continuous improvement and optimization. You will collaborate closely with development and operations teams to improve the platform stability and performance, ensuring that our systems are reliable, secure, and scalable., Infrastructure Management: * Manage and scale cloud-based infrastructure (e.g., AWS, Azure, GCP). * Apply Infrastructure as Code (IaC) principles for provisioning and configuration management. Security and Compliance: * Collaborate with the security team to implement best practices for system and data security. * Ensure systems comply with relevant industry standards and regulations. Monitoring and Performance: * Set up and maintain monitoring and alerting systems for early issue detection and resolution. * Continuously optimize system performance and resource usage. Documentation: * Create and maintain thorough documentation for SRE/platform processes, tools, and practices. Exposure to Jira and equivalent tool would be beneficial ## Related Videos - [DevOps at Netflix](https://www.wearedevelopers.com/videos/270-devops-at-netflix) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Collaboration Quantified: Lessons from Open Source Developer Networks](https://www.wearedevelopers.com/videos/1422-collaboration-quantified-lessons-from-open-source-developer-networks) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [Applying Agile Principles to Incident Management ](https://www.wearedevelopers.com/videos/101-applying-agile-principles-to-incident-management) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)