IT Service Resilience Manager

A&o Shearman
Belfast, UK
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
£83,124.0
Working hours
Regular working hours

Tech stack

Agile Methodology Software as a Service Cloud Computing Cyber Security DevOps Disaster Recovery Middleware Fault Tolerance Information Technology Operations Cloud Services Data Streaming Cybercrime

Job description

The successful candidate should have strong technical hands-on skills in Cloud, infrastructure and applications deployments, along with the ability to translate business continuity and availability requirements into technical enterprise architecture along with operational disaster recovery requirements. It will be their responsibility to own the program of resilience and recovery testing across applications, SaaS and third-party providers and internal teams.

Ownership and leadership execution in the following areas:

  • Develop and maintain dependency maps that capture application, middleware, cloud services, data flows and third-party dependencies to identify single points of failure and inform resilience design

  • Lead DR implementation and testing program: design automated DR processes when feasible, schedule regular tests

  • Own operational runbooks, monitoring and incident playbooks aligned to graceful degradation modes; ensure monitoring and SRE/operations practices are aligned to expected degradation behaviors.

  • Coordinate crisis response governance and periodic scenario exercises with crisis response teams, define activation criteria, maintain war-room procedures and ensure lessons-learned feed back into architecture and DR plans.

  • Run supplier resilience assessments for critical SaaS/third parties using a posture assessment approach; escalate remediation, negotiate contractual improvements or recommend contingencies/alternative sourcing.

  • Develop and maintain dependency maps that capture application, cloud services, data flows and third-party dependencies to identify single points of failure and inform resilience design.

  • Collaborate with enterprise architecture, security/CISO, application owners, BC/operational leads and procurement to embed resilience standards across lifecycle

  • Manage and test application/service tiers to business-agreed RTO/RPO and reliability design targets

  • Define and ensure adherence to DR/Resilience programme metrics: frequency of tests, % successful automated DR runs, closure rate for remediation actions identified through testing

  • Manage vendor performance and contractual compliance of vendors agreed operational SLAs and vendor contingency plans validated via tests.

  • Identify and assess IT resilience risks related to system outages, cyber threats and 3rd party dependencies

Requirements

  • 10+ years in technology resilience, disaster recovery, or IT operations, with 5+ years in leadership positions managing cross-functional teams.

  • Deep hands-on knowledge of a range of IT environments, SaaS, cloud infrastructure (AWS & Azure), and security tools.

  • Required expertise in ISO 22301, NIST and ITIL

  • Certifications (Preferred): Certified Business Continuity Professional (CBCP), CISSP, CISM, or DRI International certifications

  • Experience of communicating to senior stakeholders and interpreting complex technical solutions to simple language.

  • Exposure of working in both Agile and Waterfall delivery methodologies.

Personal

  • Ability to anticipate risks and shift from reactive disaster recovery to proactive service resilience, focusing on “prevention by design”

  • Skilled at navigating changing technology environments (e.g. cloud, DevOps) and leading transformation
  • Strong stakeholder engagement and influence skills to work with EA, Platform Owners, I&O, InfoSec, Business Continuity, Procurement
  • Proven ability to manage crisis situations, make quick, informed decisions during incidents, and maintain confidence (strategic optimism) within teams
  • Excellent customer-facing skills with a good grasp of key drivers and requirements within the business.

  • Understanding of how technology resilience directly impacts business operations, continuity, and profitability.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.co.uk

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

1:18 min

Implementing routing middleware for seamless multi-fragment origination

Igor Minar Igor Minar +1 · WWC 2025

4:34 min

Motivational categories behind modern cybercriminal activities

Mauro Verderosa · LIVE

3:53 min

Applying software development methodologies to incident response

Tobias Dunn-Krahn · LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

3:46 min

Navigating a career in cloud transformation consulting

Piet Van Dongen · LIVE

Videos

See all

Related articles

See all