Site Reliability Engineer II - AI & Corporate Risk Tech

JPMorgan Chase & Co.
Glasgow, UK
4 days ago
Apply on jpmc.fa.oraclecloud.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Java (Programming Language) .NET Framework Artificial Intelligence Software Applications Cloud Computing Continuous Delivery Continuous Integration Python (Programming Language) Operational Data Store Reliability Engineering Data Processing Computer Network Technologies
+4 more
Spring-boot Containerization Kubernetes Programming Languages

Job description

As a Site Reliability Engineer II at JPMorganChase within Corporate Risk Technology, you will solve complex and broad business problems with simple, straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure - independently decomposing and iteratively improving on existing solutions. You are a meaningful contributor to your team, sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform., * Guide and support team members in building appropriate-level designs, gaining peer consensus, and driving adoption of site reliability engineering best practices across the team

  • Collaborate with software engineers and cross-functional teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery pipelines
  • Implement infrastructure, configuration, and network as code for the applications and platforms within your scope
  • Partner with technical experts, key stakeholders, and team members to resolve complex problems and proactively address issues using service level indicators and objectives before they impact customers
  • Identify and address roadblocks, propose improvements to solve business problems, and explore new technologies where appropriate
  • Apply familiarity with availability, reliability, and scalability principles to iteratively improve outcomes in collaboration with partners
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements
  • Applies enterprise-authorized AI capabilities within the work environment to identify patterns in operational signals that indicate reliability risk or recurring toil, prioritizing reuse-first improvements tied to service level objective outcomes

Requirements

  • Formal training or certification on site reliability engineering concepts and proficient applied experience
  • Proficiency in site reliability culture and principles, with the ability to implement site reliability practices within an application or platform
  • Proficiency in at least one programming language such as Python, Java/Spring Boot, or .NET
  • Experience in observability practices such as white and black box monitoring, service level objective alerting, and telemetry collection
  • Proficient knowledge of software applications and technical processes within a given technical discipline (e.g., cloud, AI, mobile platforms)
  • Working knowledge of using enterprise-authorized AI capabilities within the work environment to support site reliability engineering workflows, with strong validation habits and awareness of data sensitivity
  • Ability to review and validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following security and data handling requirements

Preferred qualifications, capabilities, and skills

  • Experience with continuous integration and continuous delivery tooling
  • Familiarity with container technologies and container orchestration platforms
  • Experience troubleshooting common networking technologies and issues
  • There’s nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world’s most complex and mission-critical systems.

About the company

There’s nothing more exciting than being at the center of a rapidly growing field in technology and applying your skills to drive innovation and modernize some of the world’s most complex and mission-critical systems. At JPMorganChase, you’ll be part of a team that values curiosity, collaboration, and continuous improvement - where your contributions directly shape the reliability and resilience of platforms that matter., J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world’s most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives., Our professionals in our Corporate Functions cover a diverse range of areas from finance and risk to human resources and marketing. Our corporate teams are an essential part of our company, ensuring that we’re setting our businesses, clients, customers and employees up for success.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jpmc.fa.oraclecloud.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:50 min

Electronic diagnostic software tools and factory production flashing

Denis Grahovac · World Congress 2021

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · World Congress 2022

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:05 min

Differences between RPA platforms and the .NET framework

Maria Doina Irimias · LIVE

4:04 min

Overview of Kubernetes operators and custom resource definitions

Philipp Krenn · World Congress 2022

1:06 min

Empowering site reliability engineers with integrated AI agents

Osmar Matos Osmar Matos · World Congress 2026 Europe

Videos

See all

Related articles

See all