Lead Infrastructure Engineer - Mainframe

JPMorgan Chase & Co.
Jersey City, NJ, United States
5 days ago
Apply on www.techcareers.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Component-Based Software Engineering Customer Information Control System (CICS) Databases IBM DB2 DevOps Firmware IBM Parallel Sysplex IBM Hardware Management Console Job Control Language (JCL) Tivoli Management Framework IBM Websphere Application Server
+12 more
Python (Programming Language) Mainframes Operational Data Store Reliability Engineering Software Engineering Scripting IT Architecture Information Technology Network Server Servicenow Programming Languages Control M

Job description

As a Lead Infrastructure Engineer at JPMorgan Chase as a part of our Mainframe and Mid-Range Compute Site Reliability and Engineering (SRE) team, we look first and foremost for people who are passionate to solving business problems through innovation and modern engineering practices. You’ll be required to apply your depth of knowledge and expertise to all aspects of Infrastructure Support and Software Development Lifecycle, as well as partner continuously with your many stakeholders daily to stay focused on common goals. We embrace a culture of experimentation and constantly strive for improvement and learning. You’ll work in a collaborative, trusting, thought-provoking environment - one that encourages diversity of thought and creative solutions that are in the best interests of our customers, globally., * Lead teams of technologists that provide end-to-end application or infrastructure service delivery for the successful business operations of the firm

  • Execute policies and procedures that ensure operational stability and availability
  • Monitor production environments for anomalies, address issues, and drive evolution of utilization of standard observability tools
  • Escalate and communicate issues and solutions to the business and technology stakeholders, actively participating from incident resolution to service restoration
  • Lead incident, problem, and change management in support of full stack technology systems, applications, or infrastructure
  • Ability to host and participate in bridge calls and communicate effectively to large group of individuals at all levels.
  • Responsible for administering, troubleshooting Mainframe related components.
  • Ability to work in large, collaborative teams to achieve organizational goals.
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate infrastructure analysis and design documentation, validating outputs and handling operational data according to sensitivity and security requirements.
  • Applies reuse-first, AI-assisted practices within delivery and automation routines to identify recurring issues and validate remediation options, ensuring changes are traceable/auditable and aligned to resiliency and security expectations.

Requirements

  • Formal training or certification on software engineering concepts and 5+ years applied experience
  • 10+ experience in operating and managing the operations of IBM z-Series environments
  • Demonstrated leadership of Operational and SRE Teams in a 24X7 support environment including all aspects of people management
  • Understanding of infrastructure architecture including servers, storage, network, database, and application components.
  • Extensive knowledge “Replication” technologies’ such as IBM CSM and GDPS
  • Expertise in administering z-Series and Hardware Management Console (HMC) including Firmware/Microcode upgrades.
  • Demonstrated understanding of security standards including a working knowledge of SSH protocol. & working knowledge of transaction-based systems, IMS, CICS, DB2 and WebSphere
  • Experience in managing ServiceNow including workflow/ticket/resolution management across a global environment 24X7. & demonstrated expertise in Incident/Problem/Change management process and procedures.
  • Able to troubleshoot priority incidents, facilitate blameless post-mortems and ensure permanent closure of incidents.
  • Demonstrated experience using enterprise-authorized AI capabilities within the work environment to support infrastructure engineering workflows with strong validation habits and awareness of data sensitivity.
  • Ability to review and validate AI-assisted recommendations before implementation, escalating when uncertain and ensuring outcomes align to resiliency, security, and auditability expectations.

Preferred qualifications, capabilities, and skills

  • Working knowledge in one or more general purpose programming languages and/or automation scripting
  • Practical experience with python development
  • Knowledge in Site Reliability Engineering - Design, code, test and deliver software to automate manual operational work.
  • Knowledge of multiple Batch Scheduling tools, notably CA-7, Control-M and Zeke. Exhibit a good working knowledge of JCL (Job Control Language)
  • Conversant with Netcool support in a large-scale environment.
  • Demonstrated ability to engage with IBM z-Series Engineering L3/L4/Build Teams on Architecture, Development, Stability & Continuous Improvement of Environment to advance the product’s vision and strategy to satisfy customer needs.

About the company

J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world’s most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.techcareers.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:25 min

Engineering career path from embedded systems to banking

John Woods John Woods · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

2:19 min

Orchestrating over-the-air firmware updates for vehicle modules

Denis Grahovac · World Congress 2021

3:48 min

Modernizing massive legacy mainframe and IBM i codebases

Maximilian Jesch Maximilian Jesch · World Congress 2026 Europe

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

2:20 min

Utilizing custom firmware for variable torque manipulation

Daniel Meilak Daniel Meilak +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all