> Markdown version of [/jobs/ext/3100835-senior-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3100835-senior-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer - **Company:** iManage LLC - **Location:** London, UK - **Experience:** Expert - **Salary:** £39,000.0 - £79,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Microsoft Azure, Bash Shell, Ubuntu (Operating System), Software as a Service, Cloud Computing, Computer Engineering, Continuous Integration, Debian Linux, Software Design Documents, Linux, DevOps, Electronic Design Automation, Python (Programming Language), Linux Servers, Nagios, Windows PowerShell, Reliability Engineering, Prometheus, Ruby, Software Engineering, Scripting, Cloud Platform System, Grafana, Containerization, Kubernetes, Infrastructure Automation Frameworks, Terraform, Docker, Golang, Programming Languages - **Published:** September 27, 2026 - **Apply:** https://www.adzuna.co.uk/jobs/details/5899590501 ## About the Role * Experience writing design documents, postmortems, and refactoring application code. * Built automation to reduce operational burden or developed internal SaaS tools. * Ability to advocate for SRE principles, such as SLOs versus SLAs, and introduce them effectively. * Experience in public cloud or hosted datacenter environments, with Azure and AKS preferred. * Passion for collaborative teamwork and influencing reliability best practices across teams. * Hands-on experience with Linux server stacks, with Ubuntu or Debian preferred. * Knowledge of cloud provisioning platforms, with Terraform preferred. * Exposure to configuration management tools, with Chef preferred. * Experience with containerization and clustering technologies, with Docker preferred. * Familiarity with observability and alerting tools such as Prometheus, Grafana, ELK, or EFK. * Practical experience with CI/CD pipelines and rollout strategies. * A bachelors degree, or equivalent experience, in Computer Engineering or a related field. * Proficiency in one or more programming languages such as Java, Python, or Golang. * Familiarity with scripting languages such as PowerShell, Bash, Python, or Ruby. ## Description * Eliminate toil through automation and software development. * Partner cross-functionally with application teams and internal stakeholders. * Create a modern, cloud-native platform that is resilient, cost-effective, and secure by default. * Scale cloud infrastructure to support our Kubernetes-based ecosystem. * Maintain the freshness and utility of platform services. * Improve the security posture of our products. * Design automation, orchestration, observability, and disaster readiness into our products. * Participate in production support and on-call rotations, providing senior-level guidance during critical events. * Lead incident management and post-incident retrospectives, and coach teams in these practices. * Engage in and often lead architectural discussions, reduce toil, and deliver scalable, resilient platforms. * Help scale our cloud platform and collaborate across teams to promote standardization and resiliency. * Provide guidance during complex technical decisions and high-impact events. Technologies: * Azure * Bash * CI/CD * Cloud * Debian * Docker * ELK * Golang * Grafana * Incident Management * Support * Java * Kubernetes * Linux * PowerShell * Prometheus * Python * Ruby * Security * Terraform * Ubuntu * DevOps * LESS ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [Designing UX for SRE Agents in High-Stakes Incidents](https://www.wearedevelopers.com/videos/100003-designing-ux-for-sre-agents-in-high-stakes-incidents) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Coroutine explained yet again 60 years later](https://www.wearedevelopers.com/videos/690-coroutine-explained-yet-again-60-years-later) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)