DevOps Support Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+12 more
Job description
This role sits at the heart of a major transformation programme, supporting the migration of applications from traditional infrastructure into modern AWS cloud environments while maintaining production stability. Around 70% of the position focuses on application support and production engineering, with the remaining 30% dedicated to cloud engineering and DevOps transformation. It is particularly well suited to Infrastructure Engineers who have successfully evolved into Cloud or DevOps Engineers and enjoy working in globally distributed, highly regulated environments. Responsibilities
- Support and modernise applications as they transition from traditional infrastructure into AWS cloud environments, ensuring stability and reliability throughout the migration.
- Maintain, optimise and troubleshoot CI/CD pipelines and deployment processes to enable efficient, repeatable and reliable releases.
- Design, build and enhance automation solutions that reduce manual operational effort and improve consistency across environments.
- Provide L2 production support, including incident triage, service restoration and escalation where appropriate, and contribute to L3 support when required.
- Work closely with Cyber Security teams to ensure applications, platforms and deployments comply with security controls and governance standards.
- Manage infrastructure configuration using Infrastructure as Code principles, including the use of tools such as Terraform and Ansible.
- Monitor platform health, reliability, availability and performance using tools such as Datadog and BigPanda, and implement improvements based on observed trends.
- Participate actively in incident management, including root cause analysis, post-incident reviews and continuous improvement initiatives.
- Support and refine operational runbooks and procedures to ensure consistent handling of production issues and changes.
- Collaborate with stakeholders across geographically distributed teams, providing clear updates and technical guidance on production and transformation activities.
- Apply strong troubleshooting skills to diagnose complex issues across AWS, Linux, networking, applications and databases.
- Balance day-to-day operational support responsibilities with longer-term platform modernisation and automation initiatives.
- Implement and support secure deployment and operational practices aligned with cyber security requirements.
- Contribute to large-scale cloud migration and transformation activities, including planning, testing and cutover support.
- Engage in assessment and interview processes as needed, including technical deep dives and scenario-based discussions on production support and incident management.
Requirements
- Minimum of 5 yearsâ experience in DevOps, Production Engineering, Site Reliability Engineering or Infrastructure Engineering roles.
- Proven track record supporting business-critical production environments, including L2 and L3 production support.
- Strong background in cloud engineering and automation, with hands-on experience in AWS.
- Solid Linux administration skills, including Red Hat Enterprise Linux.
- Good understanding of networking fundamentals relevant to cloud and production environments.
- experience maintaining and improving CI/CD pipelines and deployment processes.
- Proficiency in Python Scripting for automation and operational tooling.
- Proficiency in Shell Scripting for system administration and automation tasks.
- Hands-on experience with Ansible for configuration management and automation.
- experience implementing and supporting Infrastructure as Code solutions, including Terraform.
- Strong familiarity with source control and deployment automation best practices.
- experience with Datadog for monitoring, observability and alerting.
- experience with BigPanda or similar event correlation and incident management platforms.
- Demonstrated capability in incident management, including structured response and communication.
- experience working with security controls and governance frameworks in operational environments.
- Exposure to secure deployment and operational practices aligned with Cyber Security initiatives.
- Excellent English communication skills, both written and verbal, suitable for stakeholder engagement.
- Ability to operate effectively under production support pressures and tight timelines.
- experience working within globally distributed teams and collaborating across time zones.
- Strong troubleshooting mindset focused on stability, reliability and rapid issue resolution.
Additional Skills & Qualifications
- background as an Infrastructure Engineer who has transitioned into a Cloud or DevOps Engineer role.
- experience supporting Cyber Security programmes or strategic security initiatives.
- Exposure to Site Reliability Engineering (SRE) concepts and practices.
- experience working in Financial Services or other highly regulated environments.
- Involvement in large-scale cloud migration or transformation programmes.
- experience with AWS services such as EC2 and EKS in production environments.
- Familiarity with HashiCorp Vault for secrets management and secure configuration.
- experience with SQL for basic querying, troubleshooting and diagnostics.
- Strong stakeholder engagement skills, with the ability to communicate complex technical topics clearly.
- Pragmatic approach to automation and continuous improvement, focusing on high-impact changes.
- Comfortable participating in technical screening interviews, coding or Scripting assessments and deep-dive technical discussions.
- Ability to balance operational responsibilities with project work and platform modernisation efforts.
About the company
Trading as TEKsystems. Allegis Group Limited, Bracknell, RG12 1RT, United Kingdom. No. 2876353. Allegis Group Limited operates as an Employment Business and Employment Agency as set out in the Conduct of Employment Agencies and Employment Businesses Regulations 2003. TEKsystems is a company within the Allegis Group network of companies (collectively referred to as âAllegis Groupâ). Aerotek, Aston Carter, EASi, Talentis Solutions, TEKsystems, Stamford Consultants and The Stamford Group are Allegis Group brands. If you apply, your personal data will be processed as described in the Allegis Group Online Privacy Notice available at our website.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.careerboard.comGood distractions
Talks and stories from around this role â technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
The Most Popular IT Jobs on the Market
Is Software Engineering Over-Saturated?
Fully Remote Software Engineer Jobs
Top-Paying Tech Jobs (with Salaries)