Devops SRE

Test Triangle
Greater London, UK
1 day ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours

Tech stack

Automation of Tests Bash Shell Software as a Service Cloud Computing Cloud Computing Security Cloud Engineering Computer Programming Continuous Delivery Continuous Integration DevOps Domain Name System (DNS) Virtual Private Networks (VPN)
+17 more
Python (Programming Language) Linux Kernel Routing Performance Tuning Public Key Infrastructure Role-Based Access Control Reliability Engineering Prometheus Google Cloud Istio Delivery Pipeline Kubernetes Deployment Automation Azure AKS Terraform Dynatrace Golang

Job description

Our Cloud Engineering team is seeking a seasoned and passionate Senior Cloud Engineer with deep hands-on development and cloud engineering expertise. In this role, you will serve as a key technical contributor within a cloud-focused engineering team, working on one of the Group’s flagship initiatives-delivering a strategic platform on Google Cloud Platform (GCP) that enables the business to realise next-generation services aligned with the Bank’s long-term vision.

Key Responsibilities

  • Architect, implement, and maintain highly resilient, scalable, and secure Kubernetes environments on GCP.
  • Engineer and optimise Kubernetes infrastructure to support multitenant workloads, ensuring robust isolation, resource efficiency, and operational scalability.
  • Design and enforce strong security controls, including OPA Gatekeeper policies, fine-grained RBAC, mTLS enforcement, and secure service mesh configurations.
  • Build, maintain, and enhance CI/CD pipelines enabling automated testing, seamless deployments, and continuous integration.
  • Diagnose and resolve complex system-level issues related to performance, scalability, networking, and automation.
  • Collaborate with cross-functional teams to deliver cloud-native solutions aligned with engineering best practices and business goals.

Required Skills & Experience

Core Cloud & DevOps Competencies

  • Extensive experience in DevOps or Site Reliability Engineering (SRE) roles across consumer or SaaS environments.
  • Strong expertise in deploying and managing production-grade Kubernetes clusters and containerised services.
  • Hands-on experience with Kubernetes (k8s) and Containers in live, high-availability environments.
  • Proven experience designing and implementing CI/CD pipelines for automated build, test, and deployment workflows.
  • Proficiency in programming languages such as Python, Go, and Bash for automation and tooling.
  • Demonstrated ability to take ownership of engineering initiatives and drive them to successful completion.
  • Strong experience developing and managing Infrastructure as Code (IaC) using Terraform.
  • Exposure to managing the full product lifecycle of cloud-native core services.
  • Hands-on experience with GCP infrastructure and services.
  • Deep understanding of cloud networking concepts such as Hybrid Connectivity, VPN, NAT, IPAM, DNS, and routing.
  • Strong knowledge of cloud security including KMS, PKI, encryption standards, and least-privilege access principles.
  • Experience with Service Mesh technologies such as Istio or Anthos for secure service-to-service communication and observability.
  • Competence in managing Istio telemetry, sidecar injection, and enforcing mTLS.
  • Experience with Anthos Config Management, GitOps-driven provisioning, and Backstage GitOps workflows.
  • Understanding of shared Kubernetes services such as CoreDNS, cert-manager, Dynatrace, Cloudability, and Infoblox.
  • Familiarity with OPA Gatekeeper for policy enforcement and tenant isolation.

Security, Observability & Performance

  • Strong security mindset with a proven track record of designing secure, resilient cloud-native systems.
  • Experience implementing observability stacks including Prometheus, Dynatrace, and OpenTelemetry.
  • Deep understanding of Linux internals, system performance tuning, and troubleshooting.
  • Familiarity with Aqua Security for container runtime protection.

CI/CD & Automation Tooling

  • Hands-on experience with Harness CI/CD for secure and automated deployment workflows.

Professional Attributes

  • Excellent verbal, written, and interpersonal communication skills with the ability to explain complex technical concepts clearly.
  • Ability to work effectively in fast-paced, dynamic environments and adapt quickly to change.
  • Strong analytical and problem-solving abilities with a focus on delivering high-quality outcomes.

Requirements

  • Extensive experience in DevOps or Site Reliability Engineering (SRE) roles across consumer or SaaS environments.
  • Strong expertise in deploying and managing production-grade Kubernetes clusters and containerised services.
  • Hands-on experience with Kubernetes (k8s) and Containers in live, high-availability environments.
  • Proven experience designing and implementing CI/CD pipelines for automated build, test, and deployment workflows.
  • Proficiency in programming languages such as Python, Go, and Bash for automation and tooling.
  • Demonstrated ability to take ownership of engineering initiatives and drive them to successful completion.
  • Strong experience developing and managing Infrastructure as Code (IaC) using Terraform.
  • Exposure to managing the full product lifecycle of cloud-native core services.
  • Hands-on experience with GCP infrastructure and services.
  • Deep understanding of cloud networking concepts such as Hybrid Connectivity, VPN, NAT, IPAM, DNS, and routing.
  • Strong knowledge of cloud security including KMS, PKI, encryption standards, and least-privilege access principles.
  • Experience with Service Mesh technologies such as Istio or Anthos for secure service-to-service communication and observability.
  • Competence in managing Istio telemetry, sidecar injection, and enforcing mTLS.
  • Experience with Anthos Config Management, GitOps-driven provisioning, and Backstage GitOps workflows.
  • Understanding of shared Kubernetes services such as CoreDNS, cert-manager, Dynatrace, Cloudability, and Infoblox.
  • Familiarity with OPA Gatekeeper for policy enforcement and tenant isolation.

Security, Observability & Performance

  • Strong security mindset with a proven track record of designing secure, resilient cloud-native systems.
  • Experience implementing observability stacks including Prometheus, Dynatrace, and OpenTelemetry.
  • Deep understanding of Linux internals, system performance tuning, and troubleshooting.
  • Familiarity with Aqua Security for container runtime protection.

CI/CD & Automation Tooling

  • Hands-on experience with Harness CI/CD for secure and automated deployment workflows., * Excellent verbal, written, and interpersonal communication skills with the ability to explain complex technical concepts clearly.
  • Ability to work effectively in fast-paced, dynamic environments and adapt quickly to change.
  • Strong analytical and problem-solving abilities with a focus on delivering high-quality outcomes.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:53 min

Configuring dynamic proxy updates with Istio Pilot

Jan Mensch Jan Mensch · World Congress 2026 Europe

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

7:15 min

Installing Istio programmatically with bash scripts

Thomas Südbröcker · LIVE

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · World Congress 2023

Videos

See all

Related articles

See all