> Markdown version of [/jobs/ext/1456931-senior-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1456931-senior-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer - **Company:** Symphony Communication Services, LLC - **Location:** Belfast, UK - **Experience:** Expert - **Salary:** £60,000.0 - £70,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Cloud Computing, Cloud Engineering, Computer Networks, Python (Programming Language), Linux System Administration, Reliability Engineering, Ansible, Scaled Agile Framework, Software Deployment, Software Engineering, Data Logging, Google Cloud, Grafana, Kubernetes, Infrastructure Automation Frameworks, Terraform, Splunk - **Published:** July 27, 2026 - **Apply:** https://www.adzuna.co.uk/jobs/details/5816500292 ## About the Role * Strong experience with IaC. + Required: Terraform + Nice to have: Terragrunt and Ansible * Deep understanding of Kubernetes and Helm. * Extensive experience with Linux administration and troubleshooting. * Hands-on experience with GCP and AWS services and cloud-native architectures. * Experience in implementing observability solutions. * Solid understanding of networking concepts and protocols. * Ability to work both independently and collaboratively. * Ability to lead and work on projects. * Ability to multitask and adapt quickly to changing priorities. * Excellent communication and problem-solving skills. We have an international team, so good communication skills are essential. * Willingness to participate in a 24/7 on-call rotation. ## Description As a Site Reliability Engineer, you will work with Agile engineering teams to provide production insight into running and operating software at-scale in a globally distributed and highly available cloud based system. You will guide the team to consider resiliency, scalability and operability implications in the choices they make during the development cycle to help foster an ownership in production mentality ("You build it, you own it"). You will also own and develop our platform by implementing and championing GitOps principles. You will play a hands-on role helping the team meet technical, operational, schedule, and business requirements. The ideal candidate will be a systems problem solver with a passion for crafting products that deliver incredible customer experiences, have deep experience with infrastructure, operational automation, data driven metrics collection, modern platform management, and a true desire to automate it rather than do it repeatedly., * Design, build, and maintain highly scalable and reliable infrastructure using Infrastructure-as-Code (IaC). * Manage and optimize Kubernetes clusters, leveraging Helm and ArgoCD for efficient application deployment and lifecycle management. * Administer and troubleshoot Linux-based systems, ensuring their performance, security, and availability. * Work extensively with both GCP and AWS services, architecting and managing cloud-native solutions. * Implement robust observability practices, including monitoring, logging and alerting to proactively identify and resolve issues, using Splunk and Grafana. * Develop and maintain automation tools using languages such as Python and Go to streamline operations and improve efficiency. * Provide production support, responding to incidents and outages, and participating in a 24/7 on-call rotation. * Troubleshoot complex networking issues and optimize network performance. * Engage in communications across all areas of the organization. ## Related Videos - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Embracing the Hybrid Cloud: Unlocking Success with Ansible](https://www.wearedevelopers.com/videos/932-embracing-the-hybrid-cloud-unlocking-success-with-ansible) - [Easy Mode Monitoring and Logging with Shiftmon](https://www.wearedevelopers.com/videos/2114-easy-mode-monitoring-and-logging-with-shiftmon) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)