Site Reliability Engineer

Tempest Vane Partners
Greater London, UK
16 days ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Airflow Bash Shell Cloud Engineering Distributed Systems Monitoring of Systems Reliability Engineering Ansible Prometheus Software Engineering Data Streaming System Availability Grafana
+6 more
Software Troubleshooting Kubernetes Apache Kafka Puppet Terraform Golang

Job description

This opportunity sits within a technology-driven organisation where engineering is fundamental to business performance. The successful candidate will work on highly scalable infrastructure supporting large-scale research, analytics, and trading environments across an international platform.

The position offers significant exposure to modern cloud-native engineering, distributed systems, automation, and platform reliability initiatives within a fast-moving and technically sophisticated environment.

The Opportunity

The incoming engineer will play a key role in the design, improvement, and automation of critical infrastructure systems across both cloud and on-premise environments.

Working closely with infrastructure, development, and platform teams, the role will focus heavily on scalability, resilience, observability, and operational efficiency.

Core responsibilities will include:

  • Engineering and maintaining enterprise-scale Kubernetes environments
  • Supporting the evolution toward cloud-native and distributed architectures
  • Driving automation initiatives across infrastructure and operational workflows
  • Developing and promoting Infrastructure-as-Code best practices
  • Enhancing platform stability, availability, and system performance
  • Contributing to monitoring, observability, and incident response capabilities
  • Partnering with global engineering teams on platform improvements and technical delivery

Requirements

Successful candidates are likely to demonstrate experience across several of the following areas:

  • Strong scripting or software engineering capability using Python, Golang, Bash, or similar
  • Deep understanding of Kubernetes and container-based infrastructure
  • Experience with Terraform, Ansible, Puppet, or equivalent automation tooling
  • Knowledge of hybrid infrastructure and distributed systems environments
  • Familiarity with observability and monitoring technologies such as Prometheus, Grafana, ELK, or Jaeger
  • Exposure to data streaming or workflow technologies including Kafka or Airflow
  • Strong troubleshooting capability with a focus on automation and reliability engineering

About the company

Tempest Vane Partners has partnered with a globally recognised investment management business seeking an experienced Infrastructure Engineer to join a high-calibre platform engineering team in London.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:26 min

Understanding Puppeteer and its underlying architectural design

Miki Lombardi · JS Congress

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

4:47 min

Automating frontend performance metrics with Google Lighthouse

Miki Lombardi · JS Congress

12:33 min

Exploring advanced observability stacks and distributed infrastructure challenges

Pawel Piwosz · LIVE

Videos

See all

Related articles

See all