Site Reliability Engineer - Data Services

iManage LLC
Belfast, UK
2 days ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Query Performance Java (Programming Language) Microsoft Azure Bash Shell Ubuntu (Operating System) Software as a Service Configuration Management Computer Engineering Data as a Services Data Stores Relational Databases Debian Linux
+23 more
Software Design Documents Distributed Data Store Elasticsearch Python (Programming Language) Linux Servers MariaDB Nagios Windows PowerShell Reliability Engineering Prometheus Ruby Software Engineering Scripting Grafana Containerization Core Data Kubernetes Hashicorp Terraform Code Restructuring Docker Golang Programming Languages

Job description

You are an engineer, a builder, and a systems thinker. You ensure data durability, optimize query performance, and manage stateful storage upgrades. You combine technical depth with empathy, working closely with customers who hold the highest expectations for the stewardship of the world’s most sensitive data.

You elevate the people around you-acting as a subject matter expert, a mentor, and an agent of change. You focus on contributing factors rather than single root causes, value code over documentation and documentation over process, and continuously seek ways to reduce toil.

You participate in architectural and design discussions, helping shape a scalable, resilient platform that supports both our customers and our organization. You collaborate across teams to drive unified, standards-based decisions that strengthen reliability. You also take part in on-call rotations and provide expertise in observability, change management, and system scalability.

As iManage experiences rapid growth in its flagship cloud product, we’re looking for engineers who bring a beginner’s mindset, embrace complexity, and care deeply about resilience and sustainability in a cloud-native world. This role includes a strong focus on the reliability and evolution of our core data services, including MariaDB, MaxScale, and Elasticsearch. If you write code, think in systems, automate relentlessly, and are passionate about reliability and scale, we want to talk to you., * Eliminating TOIL through automation and software development.

  • Partnering productively and cross-functionally with application teams and other internal stakeholders.
  • Creating a modern, cloud-native platform that is resilient, cost effective, and secure by default.
  • Scaling and tuning high-availability data clusters (MariaDB, MaxScale, Elasticsearch) in a Kubernetes environment.
  • Maintaining the freshness and utility of our platform services.
  • Improving the security posture of our products.
  • Writing / designing automation, orchestration, observability, and disaster readiness into our products.
  • Coordinating and participating in production support and on-call rotations.
  • Leading incident management efforts and post-incident retrospectives.
  • Comfortability writing design documents / postmortems and refactoring application code when needed.
  • Experience operating or supporting distributed data systems (e.g., relational databases, search clusters, or sharded storage systems).
  • Developed automation to reduce the operational burden of a product or developed software-as-a-service for internal customers., * Join a rapidly evolving, industry-leading SaaS company on an exciting journey of growth and scalability!
  • Take on meaningful, high-impact challenges by leveraging cutting-edge technologies and best-in-class protocols to drive innovation.
  • Own my career path with our internal development framework. Ask us more about this!
  • Expand my skill set and earn certifications with unlimited access to LinkedIn Learning courses and interactive Microsoft courses & training.
  • Be part of a supportive and experienced team within a dynamic, inclusive, and encouraging culture.
  • Enjoy flexible work hours that empower me to balance personal time with professional commitments.
  • Collaborate in a modern, open-plan workspace featuring a gaming area, free snacks and drinks, and regular social events.

Requirements

  • Ability to advocate for SRE concepts such as Google’s SRE concepts (e.g., I know the differences between an SLO and an SLA and can effectively introduce them to an organization).
  • Experience working in a public cloud and/or hosted datacenter environment (Azure and AKS strongly preferred).
  • A passion for working collaboratively with other teams., * Comfortability writing design documents / postmortems and refactoring application code when needed.
  • Experience operating or supporting distributed data systems (e.g., relational databases, search clusters, or sharded storage systems).
  • Developed automation to reduce the operational burden of a product or developed software-as-a-service for internal customers.
  • Ability to advocate for SRE concepts such as Google’s SRE concepts (e.g., I know the differences between an SLO and an SLA and can effectively introduce them to an organization).
  • Experience working in a public cloud and/or hosted datacenter environment (Azure and AKS strongly preferred).
  • A passion for working collaboratively with other teams., * Hands on experience with MariaDB, MaxScale, or Elasticsearch in production environments.
  • Familiarity with data store observability, query performance tuning, or capacity planning.
  • Hands on experience with Linux Server stacks (Ubuntu/Debian distributions preferred).
  • Knowledge of cloud provisioning platforms (HashiCorp Terraform preferred).
  • Exposure to at least one configuration management platform (Chef preferred).
  • Experience with containerization/clustering technologies (Docker preferred).
  • Comfortability with observability and alerting tools (Prometheus/Grafana or ELK/EFK preferred).
  • Practical experience with CI/CD pipelines and ability to describe the pros/cons associated with different rollout strategies.
  • A bachelor’s degree (or equivalent experience) in Computer Engineering or a related field.
  • Demonstrable proficiency in one or more programming languages (e.g., Java, Python, Golang).
  • Familiarity with at least one scripting language (e.g., PowerShell, Bash, Python, Ruby).

Benefits & conditions

  • Creating an inclusive environment where you’re encouraged to help shape the culture by bringing your unique perspective, not just by fitting in.
  • Providing a market-leading salary determined through a fair and consistent process, equitable for all our employees, and regularly reviewed against industry benchmarks.
  • Rewarding me with an annual performance-based bonus.
  • Providing enhanced parental leave (20 weeks for primary and 10 weeks for secondary caregiver at 100% pay)
  • Matching my pension contribution (up to 6%)
  • Offering BUPA private medical insurance & a Simplyhealth cash plan to assist with the everyday costs.
  • Providing Group life cover, including life insurance, income protection, and critical illness protection.
  • Encouraging me to make use of our top-tier flexible time off policy, which includes 25 days of annual leave and the flexibility to take further additional time off as needed
  • Having multiple company wellness days each year to prioritize mental health and well-being.
  • Providing access to RethinkCare, a global behavioral health platform that enhances personal well-being, strengthens professional resilience, and empowers parental success through expert-led training and resources.

About the company

At iManage, we are dedicated to Making Knowledge Work . Our intelligent, cloud-enabled, and secure platform is trusted by 4,100+ customers and 430,000 users worldwide, managing over 11 billion documents and 11 petabytes of data. We empower professionals across 65+ countries to unlock the full potential of their business content and communications.

We are continuously innovating to solve the most complex professional challenges and enable better business outcomes; Our work is not always easy but it is ambitious and rewarding.

So we’re looking for people who embracechallenges. People who thrive on solving problems, pushing boundaries, and collaborating with the industry’s best and brightest. That’s the iManage way. It’s how we turn the impossible into reality, empower our employees to grow, unlock their potential, and create a meaningful impact on everything we do.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

3:37 min

Why differing legacy workflows complicate monitoring tool migrations

Mathias Palmersheim Mathias Palmersheim · Europe 2026 Virtual

50 sec

Why developer happiness matters in web frameworks

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

6:16 min

Event-driven Golang backend architecture and cloud deployment

Irina Branovic Irina Branovic · World Congress 2026 Europe

Videos

See all

Related articles

See all