> Markdown version of [/jobs/ext/3296159-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3296159-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Delinea - **Location:** Belfast, UK (Remote available) - **Experience:** Expert - **Salary:** £70,000.0 - £85,000.0 - **Contract:** Permanent contract - **Skills:** Microsoft Azure, Software as a Service, Cloud Computing, Cloud Computing Security, Continuous Integration, DevOps, Disaster Recovery, Domain Name System (DNS), Failover, Virtual Private Networks (VPN), Python (Programming Language), Network Troubleshooting, Log Analysis, Networking Basics, Routing, Windows PowerShell, Redis, Reliability Engineering, SQL Databases, Datadog, Data Logging, Load Balancing, Cloud Platform System, Cloud Monitoring, Multi-Cloud, Break Fix, Firewalls (Computer Science), Kubernetes, Infrastructure Automation Frameworks, Terraform, Elk Stack - **Published:** September 3, 2026 - **Apply:** https://uk.indeed.com/viewjob?jk=d89db3b87752bc14 ## About the Role * 5+ years of relevant experience in Site Reliability Engineering, DevOps, or Cloud Administration, with demonstrated ownership of production systems. * Hands-on experience administering Azure environments, including AKS (Kubernetes), core Azure services, cloud networking, and cloud security fundamentals. * Solid understanding of monitoring, logging, and alerting practices (e.g., Datadog, Azure Monitor, ELK stack), including hands-on troubleshooting with log analysis and stack traces using Datadog APM. * Familiarity with networking fundamentals: firewalls, load balancers, VPNs, DNS, and routing. * Experience with automation and scripting (PowerShell, Python, or similar). * Practical understanding of backup, redundancy, and disaster recovery strategies in cloud environments, including geo-redundant / multi-region deployments. * Strong ownership mindset across the full incident lifecycle, from detection through post-mortem, with a customer-first approach. * Comfort communicating clearly and professionally in written form, including incident status updates and post-incident summaries. We'd Love to See: * Experience with AWS Cloud Platform. * Experience with CI/CD tools such as Azure DevOps. * Experience with infrastructure-as-code tools such as Terraform or ARM templates. * Prior experience operating SaaS products with regional tenant architectures (e.g., multiple geo-specific production environments). ## Description Our growing technology company is seeking an experienced Site Reliability Engineer with deep Azure expertise to help maintain the availability, performance, and reliability of our critical SaaS applications. In this role, you will own and drive automation, monitoring, incident response, and infrastructure improvements across our multi-cloud, multi-region environment, working closely with senior engineering and cross-functional teams., * Own the availability and performance of production SaaS applications running on Azure (AKS, App Service, Redis, SQL, Service Bus etc), across multiple geographic regions. * Lead troubleshooting and resolution of cloud infrastructure and application issues, including AKS pod/node failures, deployment rollbacks, ingress and networking issues, and resource/autoscaling problems. * Participate in an on-call rotation (including weekends) and drive incident response from detection through resolution, with a primary focus on customer experience and minimizing impact. * Drive improvements to disaster recovery, failover, and incident management processes across multi-region deployments. * Build and maintain automation scripts and monitoring tools to reduce manual toil and streamline operational tasks. * Author post-incident reviews (RCAs), identify root causes, and drive preventive action items to closure. * Partner with senior engineers and cross-functional teams to implement and improve reliability, observability, and performance best practices. * Contribute to continuous improvement initiatives across infrastructure, tooling, and process. * Communicate clearly with customer-facing stakeholders when incidents require external status updates or written incident summaries. ## Related Videos - [Azure-Well Architected Framework - designing mission critical workloads in practice](https://www.wearedevelopers.com/videos/1529-azure-well-architected-framework-designing-mission-critical-workloads-in-practice) - [The Memory Leak That Ate Our Cluster: A Postmortem](https://www.wearedevelopers.com/videos/2057-the-memory-leak-that-ate-our-cluster-a-postmortem) - [Shifting Stress to Progress— Understanding DevOps to do DevOps Better](https://www.wearedevelopers.com/videos/268-shifting-stress-to-progress-understanding-devops-to-do-devops-better) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Accelerating Authentication Architecture: Taking Passwordless to the Next Level](https://www.wearedevelopers.com/videos/733-accelerating-authentication-architecture-taking-passwordless-to-the-next-level) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) ## Related Articles - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)