Infrastructure Engineer (Storage, WEKA/CEPH): £200k + Bonus

Hunter Bond
Greater London, UK
15 days ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
£200,000.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Amazon S3 Cloud Storage Configuration Management Continuous Integration Data Integrity Linux DevOps Disaster Recovery Disk Arrays Distributed Data Store Elasticsearch
+19 more
General Parallel File Systems Information Lifecycle Management Storage Area Network (SAN) Python (Programming Language) Network File Systems Ansible Prometheus Runbook Weka Ceph (Software) Datadog High Performance Computing Grafana Software Troubleshooting Storage Technologies Data Management Kibana Software Version Control Nvme

Job description

We are partnering with a leading global investment and technology-driven organisation seeking an experienced Platform Storage Specialist to join its growing infrastructure team in London.

This is an opportunity to work on large-scale, business-critical storage platforms supporting high-performance compute environments across a global estate. The successful candidate will play a key role in designing, building, automating, and optimising enterprise storage services that underpin a rapidly expanding technology environment.

Key Responsibilities

  • Design, deploy, and manage distributed storage platforms at scale
  • Architect and maintain high-performance storage solutions including NFS, GPFS, WEKA, object storage, and related technologies
  • Perform capacity planning, forecasting, and lifecycle management across storage environments
  • Monitor and optimise storage performance, identifying bottlenecks and implementing improvements
  • Develop automation and tooling for storage provisioning, configuration management, and operational efficiency
  • Collaborate closely with infrastructure, networking, and engineering teams to deliver highly available storage services
  • Enhance monitoring, observability, alerting, and incident response capabilities
  • Evaluate and implement emerging storage technologies and cloud storage services
  • Ensure data integrity, resilience, and disaster recovery readiness across critical platforms
  • Produce and maintain technical documentation, operational procedures, and runbooks

Required Experience

  • 5+ years of experience in infrastructure engineering with a strong focus on Linux and enterprise storage administration
  • Deep knowledge of distributed storage technologies such as GPFS, WEKA, Lustre, Ceph, NFS, or similar platforms
  • Experience with storage performance analysis, benchmarking, and optimisation
  • Strong automation skills using Python and infrastructure management tools such as Ansible or Chef
  • Experience working with CI/CD pipelines, version control systems, and modern engineering practices
  • Knowledge of observability platforms such as Prometheus, Grafana, Elasticsearch, Kibana, or Datadog
  • Experience supporting cloud storage services across AWS and/or GCP
  • Strong troubleshooting, communication, and stakeholder management skills
  • Object storage platforms such as MinIO, Cloudian, or S3-compatible technologies
  • Storage hardware, NVMe, SSD, disk arrays, storage networking, and HBAs
  • Backup, replication, disaster recovery, and data protection strategies
  • High-performance computing (HPC) environments
  • Storage tiering and data lifecycle management

This position offers the opportunity to work within a highly technical environment where infrastructure performance, scalability, and automation are core to the organisation’s success.

If interested, please apply with your updated CV.

Requirements

  • 5+ years of experience in infrastructure engineering with a strong focus on Linux and enterprise storage administration
  • Deep knowledge of distributed storage technologies such as GPFS, WEKA, Lustre, Ceph, NFS, or similar platforms
  • Experience with storage performance analysis, benchmarking, and optimisation
  • Strong automation skills using Python and infrastructure management tools such as Ansible or Chef
  • Experience working with CI/CD pipelines, version control systems, and modern engineering practices
  • Knowledge of observability platforms such as Prometheus, Grafana, Elasticsearch, Kibana, or Datadog
  • Experience supporting cloud storage services across AWS and/or GCP
  • Strong troubleshooting, communication, and stakeholder management skills
  • Object storage platforms such as MinIO, Cloudian, or S3-compatible technologies
  • Storage hardware, NVMe, SSD, disk arrays, storage networking, and HBAs
  • Backup, replication, disaster recovery, and data protection strategies
  • High-performance computing (HPC) environments
  • Storage tiering and data lifecycle management

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

2:14 min

Solving complex platform architecture challenges at an enterprise scale

Maria Apazoglou · Coffee With Developers

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

8:22 min

Simulating a Linux terminal and running Spring Boot

Jakov Semenski · LIVE

Videos

See all

Related articles

See all