Sr Lead Infrastructure Engineer- Devops/AWS

JPMorganChase
Glasgow, UK
7 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Application Release Automation Cloud Engineering Continuous Integration DevOps Python (Programming Language) Key Management Reliability Engineering Azure Machine Learning Software Engineering Datadog
+9 more
Data Logging Scripting Large Language Models Containerization Kubernetes Infrastructure Automation Frameworks Information Technology Machine Learning Operations Terraform

Job description

This is a Vice President-level role and an integral part of the IPB Tech AIML team, reporting to the Head of AI, IPB Tech., * Owns and evolves the team’s CI/CD pipelines, release automation, and deployment tooling

  • Establishes reliability practices (SLOs, error budgets, runbooks) and leads production incident response and post-incident review
  • Builds and operates observability across the team’s AI/ML services (metrics, logging, tracing, alerting)
  • Automates infrastructure provisioning and configuration through infrastructure-as-code
  • Implements operational security, secrets management, and access controls to firm-wide standards
  • Partners with platform, data, and AI engineers to harden services for production and reduce dependency on external functions
  • Mentors junior engineers on DevOps and reliability practices and sets standards through review
  • Champions the firm’s culture of diversity, Opportunity, inclusion, and respect

Requirements

  • Formal training or certification on software engineering or systems concepts and applied experience
  • Advanced proficiency with infrastructure-as-code (e.g., Terraform) and scripting in Python and/or shell
  • Deep hands-on experience with CI/CD tooling and building release automation at scale
  • Strong experience with Kubernetes, containerisation, and cloud-native operations
  • Proven experience running production services: observability, on-call, incident response, and reliability engineering
  • Understanding of production security and change-management controls
  • Strong communication skills and the ability to set operational standards across a team
  • Formal SRE experience in a regulated or high-availability environment
  • Master’s degree in Computer Science, Engineering, or a related technical field (or equivalent applied experience), * Experience operating ML / LLM workloads in production (MLOps, inference reliability, cost/performance management)
  • Experience within financial services technology
  • Familiarity with JPM-internal platform, cloud, and observability tooling for internal candidates

About the company

J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world’s most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives., J.P. Morgan Asset & Wealth Management delivers industry-leading investment management and private banking solutions. Asset Management provides individuals, advisors and institutions with strategies and expertise that span the full spectrum of asset classes through our global network of investment professionals. Wealth Management helps individuals, families and foundations take a more intentional approach to their wealth or finances to better define, focus and realize their goals.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

1:36 min

Visualizing memory limits and isolating suspicious endpoints

Dina Matveev Dina Matveev · Europe 2026 Virtual

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

5:39 min

Transitioning from technical engineering into corporate project leadership

Peter Busch · LIVE

Videos

See all

Related articles

See all