Site Reliability Engineer (Hyper-V Infrastructure)

Uniting People
London, UK
5 days ago

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
4 years minimum
Compensation
£117,000.0
Working hours
Regular working hours

Tech stack

Microsoft Azure Backup Devices Disaster Recovery Failover Clustering Hyper-V System Center Operations Management Windows Servers Windows PowerShell Reliability Engineering Ansible Prometheus Runbook
+14 more
System Center Virtual Machine Management Virtual Machine Manager Private Cloud Environment Cloud Monitoring System Availability Grafana HybridCloud Infrastructure Automation Frameworks Storage Technologies Information Technology Performance Monitor Veeam Terraform Splunk

Job description

We’re recruiting for an experienced Site Reliability Engineer (SRE)/Infrastructure Engineer to join a major Banking organisation, supporting a mission-critical Microsoft Hyper-V Private Cloud platform.

This is a hands-on engineering role focused on maintaining the reliability, availability and performance of a large-scale Hyper-V estate. You’ll work within a highly regulated environment, delivering production support, infrastructure automation, platform upgrades and continuous service improvements.

This opportunity is ideal for an Infrastructure Engineer with strong Hyper-V, Windows Server and PowerShell expertise who enjoys solving complex production issues and improving operational resilience., * Administer and support enterprise Microsoft Hyper-V infrastructure.

  • Manage Hyper-V Failover Clusters, storage, networking and high availability.
  • Perform Windows Server administration, patching, upgrades and life cycle management.
  • Support private cloud and VDI platforms within a production environment.
  • Automate operational tasks using PowerShell.
  • Manage disaster recovery, backup and business continuity activities.
  • Perform proactive monitoring, incident management, root cause analysis and service improvement.
  • Support infrastructure migrations including P2V, V2V and platform modernisation.
  • Work closely with Infrastructure, Security and Platform Engineering teams to improve reliability and operational efficiency.
  • Produce technical documentation, runbooks and operational procedures.

Requirements

Essential Skills (must have)

  • Strong Microsoft Hyper-V administration experience.
  • Windows Server 2016/2019/2022 administration.
  • Hyper-V Failover Clustering and High Availability.
  • PowerShell Scripting and infrastructure automation.
  • Storage technologies including SAN/NAS, Storage Spaces Direct (S2D) or Cluster Shared Volumes (CSV).
  • SCVMM (System Center Virtual Machine Manager).
  • Disaster Recovery, Backup and Hyper-V Replica.
  • Experience supporting enterprise production infrastructure.
  • Strong troubleshooting and incident management skills.
  • ITIL environment experience
  • Experience working within Banking or Financial Services.

Desirable Skills

  • Azure Monitor, SCOM, Splunk, Grafana or Prometheus.
  • Infrastructure as Code (Terraform, Ansible or similar).
  • VDI technologies.
  • Azure or Hybrid Cloud.
  • Veeam Backup.

Experience Required

  • 6+ years’ Infrastructure Engineering experience.
  • 4+ years’ hands-on Microsoft Hyper-V administration.
  • Experience supporting highly available production environments.
  • Strong understanding of operational resilience, platform reliability and continuous service improvement.
  • Previous experience within a regulated Financial Services or Banking environment is highly desirable.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerboard.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

3:28 min

Recognizing vital enterprise stability in legacy software development roles

Gunnar Grosch · Coffee With Developers

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

4:19 min

Introduction to network security and endpoint monitoring architectures

Christoph Ruggenthaler · LIVE

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

Videos

See all

Related articles

See all