Senior Site Reliability Engineer

iManage LLC
London, UK
1 day ago
Apply on www.adzuna.co.uk
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
£39,000.0 - £79,000.0
Working hours
Regular working hours

Tech stack

Java (Programming Language) Microsoft Azure Bash Shell Ubuntu (Operating System) Software as a Service Cloud Computing Computer Engineering Continuous Integration Debian Linux Software Design Documents Linux DevOps
+19 more
Electronic Design Automation Python (Programming Language) Linux Servers Nagios Windows PowerShell Reliability Engineering Prometheus Ruby Software Engineering Scripting Cloud Platform System Grafana Containerization Kubernetes Infrastructure Automation Frameworks Terraform Docker Golang Programming Languages

Job description

  • Eliminate toil through automation and software development.
  • Partner cross-functionally with application teams and internal stakeholders.
  • Create a modern, cloud-native platform that is resilient, cost-effective, and secure by default.
  • Scale cloud infrastructure to support our Kubernetes-based ecosystem.
  • Maintain the freshness and utility of platform services.
  • Improve the security posture of our products.
  • Design automation, orchestration, observability, and disaster readiness into our products.
  • Participate in production support and on-call rotations, providing senior-level guidance during critical events.
  • Lead incident management and post-incident retrospectives, and coach teams in these practices.
  • Engage in and often lead architectural discussions, reduce toil, and deliver scalable, resilient platforms.
  • Help scale our cloud platform and collaborate across teams to promote standardization and resiliency.
  • Provide guidance during complex technical decisions and high-impact events.

Technologies:

  • Azure
  • Bash
  • CI/CD
  • Cloud
  • Debian
  • Docker
  • ELK
  • Golang
  • Grafana
  • Incident Management
  • Support
  • Java
  • Kubernetes
  • Linux
  • PowerShell
  • Prometheus
  • Python
  • Ruby
  • Security
  • Terraform
  • Ubuntu
  • DevOps
  • LESS

Requirements

  • Experience writing design documents, postmortems, and refactoring application code.
  • Built automation to reduce operational burden or developed internal SaaS tools.
  • Ability to advocate for SRE principles, such as SLOs versus SLAs, and introduce them effectively.
  • Experience in public cloud or hosted datacenter environments, with Azure and AKS preferred.
  • Passion for collaborative teamwork and influencing reliability best practices across teams.
  • Hands-on experience with Linux server stacks, with Ubuntu or Debian preferred.
  • Knowledge of cloud provisioning platforms, with Terraform preferred.
  • Exposure to configuration management tools, with Chef preferred.
  • Experience with containerization and clustering technologies, with Docker preferred.
  • Familiarity with observability and alerting tools such as Prometheus, Grafana, ELK, or EFK.
  • Practical experience with CI/CD pipelines and rollout strategies.
  • A bachelors degree, or equivalent experience, in Computer Engineering or a related field.
  • Proficiency in one or more programming languages such as Java, Python, or Golang.
  • Familiarity with scripting languages such as PowerShell, Bash, Python, or Ruby.

About the company

We are iManage, a global organization dedicated to Making Knowledge Work with an intelligent, cloud-enabled, secure platform trusted by 4,100+ customers and 430,000 users worldwide, managing over 11 billion documents and 11 petabytes of data. We work in distributed teams anchored to offices around the globe, with Tuesdays and Thursdays dedicated to in-office collaboration and Mondays and Fridays reserved for remote-friendly focus time. We offer a supportive, inclusive culture with flexible work hours, a modern open-plan workspace with a gaming area, free snacks and drinks, regular social events, internal development opportunities, unlimited access to LinkedIn Learning and Microsoft training, market-leading salary, annual bonus, enhanced parental leave, pension matching, private medical insurance, life cover, flexible time off, wellness days, and access to RethinkCare. This role is based in London, United Kingdom, within our Cloud Operations department, and is full time for an

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.co.uk
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:37 min

Why differing legacy workflows complicate monitoring tool migrations

Mathias Palmersheim Mathias Palmersheim · Europe 2026 Virtual

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

50 sec

Why developer happiness matters in web frameworks

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

8:02 min

Integrating service level objectives into incident management

Diana Todea · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

Videos

See all

Related articles

See all