Lead Site Reliability Engineer

Ocho People
Belfast, UK
13 days ago
Apply on www.adzuna.co.uk
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
£100,000.0 - £120,000.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Software System Penetration Testing Linux Elasticsearch MongoDB Reliability Engineering Ansible Data Ingestion Kubernetes Apache Kafka Vertica Docker

Job description

Ocho are working with a client to find a Lead Site Reliability Engineer (SRE) to lead the team responsible for keeping their platform running reliably and securely, 24/7.

Our client helps thousands of teams in 60+ countries monitor and improve their applications, and is remote-first, valuing impact, transparency and continuous improvement.

The role

This is a player-coach position. You’ll set the technical direction and own reliability and security strategy for the platform, while staying hands-on with the systems your team runs. It’s a small team with a long-standing habit of fixing root causes, not just alerts, and they’re now growing it as the business scales.

Their stack

  • Mostly bare-metal infrastructure, managed by Ansible
  • Data ingestion and processing in Rust, running on Kafka
  • A Rails app serving the customer-facing UI
  • MongoDB, ClickHouse and ElasticSearch

Responsibilities

  • Lead the SRE team: set priorities, mentor engineers, grow the team
  • Own reliability strategy and the long-term infrastructure roadmap
  • Be part of the on-call rotation, and keep improving it
  • Act as incident coordinator, and lead blameless postmortems
  • Guide strategic projects, including new AWS infrastructure
  • Stay hands-on: tune the Rust codebase and infrastructure automation
  • Handle security researcher reports, coordinate penetration tests, support ISO renewals

Requirements

  • 8+ years keeping large Linux systems reliable, with experience leading an SRE, platform or infrastructure team (formally or as a technical lead).
  • Competent developer across multiple languages, ideally with Rust and Ansible experience.
  • Strong incident response and postmortem experience, comfortable translating business growth into infrastructure strategy.
  • Bonus: AWS, Kubernetes and Docker.

Benefits & conditions

  • Competitive salary
  • Remote-first culture - UK wide
  • Stock options,
  • Flexible PTO
  • Personal development budget.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.co.uk
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:01 min

Migrating existing applications from MongoDB to Postgres

Nikita Shamgunov Nikita Shamgunov · World Congress 2024

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

6:58 min

Building engineering communities and finding technical inspiration

Videos

See all

Related articles

See all