DevOps Support Engineer

TEKsystems
London, UK
7 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
ÂŁ117,000.0 - ÂŁ169,000.0
Working hours
Regular working hours
Languages
English

Tech stack

Amazon Web Services Amazon Elastic Compute Cloud Cloud Computing Cloud Engineering Configuration Management Cyber Security Databases Linux DevOps Key Management Linux System Administration Networking Basics
+12 more
Red Hat Enterprise Linux Reliability Engineering Ansible Shell Script SQL Databases Datadog Scripting Software Security Deployment Automation Hashicorp Terraform Software Version Control

Job description

This role sits at the heart of a major transformation programme, supporting the migration of applications from traditional infrastructure into modern AWS cloud environments while maintaining production stability. Around 70% of the position focuses on application support and production engineering, with the remaining 30% dedicated to cloud engineering and DevOps transformation. It is particularly well suited to Infrastructure Engineers who have successfully evolved into Cloud or DevOps Engineers and enjoy working in globally distributed, highly regulated environments. Responsibilities

  • Support and modernise applications as they transition from traditional infrastructure into AWS cloud environments, ensuring stability and reliability throughout the migration.
  • Maintain, optimise and troubleshoot CI/CD pipelines and deployment processes to enable efficient, repeatable and reliable releases.
  • Design, build and enhance automation solutions that reduce manual operational effort and improve consistency across environments.
  • Provide L2 production support, including incident triage, service restoration and escalation where appropriate, and contribute to L3 support when required.
  • Work closely with Cyber Security teams to ensure applications, platforms and deployments comply with security controls and governance standards.
  • Manage infrastructure configuration using Infrastructure as Code principles, including the use of tools such as Terraform and Ansible.
  • Monitor platform health, reliability, availability and performance using tools such as Datadog and BigPanda, and implement improvements based on observed trends.
  • Participate actively in incident management, including root cause analysis, post-incident reviews and continuous improvement initiatives.
  • Support and refine operational runbooks and procedures to ensure consistent handling of production issues and changes.
  • Collaborate with stakeholders across geographically distributed teams, providing clear updates and technical guidance on production and transformation activities.
  • Apply strong troubleshooting skills to diagnose complex issues across AWS, Linux, networking, applications and databases.
  • Balance day-to-day operational support responsibilities with longer-term platform modernisation and automation initiatives.
  • Implement and support secure deployment and operational practices aligned with cyber security requirements.
  • Contribute to large-scale cloud migration and transformation activities, including planning, testing and cutover support.
  • Engage in assessment and interview processes as needed, including technical deep dives and scenario-based discussions on production support and incident management.

Requirements

  • Minimum of 5 years’ experience in DevOps, Production Engineering, Site Reliability Engineering or Infrastructure Engineering roles.
  • Proven track record supporting business-critical production environments, including L2 and L3 production support.
  • Strong background in cloud engineering and automation, with hands-on experience in AWS.
  • Solid Linux administration skills, including Red Hat Enterprise Linux.
  • Good understanding of networking fundamentals relevant to cloud and production environments.
  • experience maintaining and improving CI/CD pipelines and deployment processes.
  • Proficiency in Python Scripting for automation and operational tooling.
  • Proficiency in Shell Scripting for system administration and automation tasks.
  • Hands-on experience with Ansible for configuration management and automation.
  • experience implementing and supporting Infrastructure as Code solutions, including Terraform.
  • Strong familiarity with source control and deployment automation best practices.
  • experience with Datadog for monitoring, observability and alerting.
  • experience with BigPanda or similar event correlation and incident management platforms.
  • Demonstrated capability in incident management, including structured response and communication.
  • experience working with security controls and governance frameworks in operational environments.
  • Exposure to secure deployment and operational practices aligned with Cyber Security initiatives.
  • Excellent English communication skills, both written and verbal, suitable for stakeholder engagement.
  • Ability to operate effectively under production support pressures and tight timelines.
  • experience working within globally distributed teams and collaborating across time zones.
  • Strong troubleshooting mindset focused on stability, reliability and rapid issue resolution.

Additional Skills & Qualifications

  • background as an Infrastructure Engineer who has transitioned into a Cloud or DevOps Engineer role.
  • experience supporting Cyber Security programmes or strategic security initiatives.
  • Exposure to Site Reliability Engineering (SRE) concepts and practices.
  • experience working in Financial Services or other highly regulated environments.
  • Involvement in large-scale cloud migration or transformation programmes.
  • experience with AWS services such as EC2 and EKS in production environments.
  • Familiarity with HashiCorp Vault for secrets management and secure configuration.
  • experience with SQL for basic querying, troubleshooting and diagnostics.
  • Strong stakeholder engagement skills, with the ability to communicate complex technical topics clearly.
  • Pragmatic approach to automation and continuous improvement, focusing on high-impact changes.
  • Comfortable participating in technical screening interviews, coding or Scripting assessments and deep-dive technical discussions.
  • Ability to balance operational responsibilities with project work and platform modernisation efforts.

About the company

Trading as TEKsystems. Allegis Group Limited, Bracknell, RG12 1RT, United Kingdom. No. 2876353. Allegis Group Limited operates as an Employment Business and Employment Agency as set out in the Conduct of Employment Agencies and Employment Businesses Regulations 2003. TEKsystems is a company within the Allegis Group network of companies (collectively referred to as “Allegis Group”). Aerotek, Aston Carter, EASi, Talentis Solutions, TEKsystems, Stamford Consultants and The Stamford Group are Allegis Group brands. If you apply, your personal data will be processed as described in the Allegis Group Online Privacy Notice available at our website.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerboard.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:32 min

Shifting to a DevOps career from non-technical backgrounds

Megha Kadur ¡ LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard ¡ WWC 2025

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum ¡ WWC Europe 2026

5:34 min

Managing token budgets and enterprise usage of coding agents

Chris Heilmann +2 ¡ LIVE

2:39 min

Experiencing core Linux capabilities for DevOps administration

Michael Cade ¡ LIVE

2:30 min

Establishing shared definitions for devops and cloud architectures

Bruno Amaro Almeida ¡ WWC 2022

Videos

See all

Related articles

See all