Associate - Infrastructure Engineer III
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+33 more
Job description
As an Infrastructure Engineer III at JPMorganChase within Enterprise Technology (Infrastructure Platforms), you apply strong knowledge of software, applications, and technical processes within the infrastructure engineering discipline. In this role, you will hold end-to-end accountability for the administration, stability, and resilience of the storage technology estate - spanning reactive incident management through to proactive automation and toil reduction. You will operate on a structured shift rotation, including weekend day coverage, to ensure continuity of service and operational excellence across the infrastructure landscape., * Apply technical knowledge and problem-solving methodologies to storage infrastructure projects of moderate scope, ensuring end-to-end monitoring, performance, and resilience of storage services running at scale
- Use enterprise-authorized AI capabilities to accelerate infrastructure analysis, monitoring, and capacity documentation, validating outputs and handling operational data according to sensitivity and security requirements
- Apply reuse-first, AI-assisted practices within delivery and automation routines to identify recurring issues, improve remediation workflows, and ensure changes are traceable, auditable, and aligned to resiliency and security expectations
- Operate and enhance block, file, and object storage platforms across on-premises and cloud environments, including performance tuning, capacity planning, lifecycle management, and resiliency testing such as failover and disaster recovery validation
- Lead incident response for storage outages and performance degradations, drive root cause analyses, and implement preventative actions to reduce recurrence
- Build and maintain automation for provisioning, patching, upgrades, replication, backup and restore, and compliance checks to reduce toil and improve operational consistency
- Create and maintain runbooks, escalation paths, and standardized operational procedures to support on-call readiness and team knowledge continuity
- Partner with infrastructure, network, operating system, database, and application teams to meet workload requirements and reliability targets
- Implement AI-driven observability and AIOps capabilities - including telemetry correlation, anomaly and regression detection, and large language model-assisted incident and runbook workflows - with a focus on accuracy, auditability, and safe rollout
- Own and continuously improve service level objectives, service level indicators, error budgets, and on-call readiness for storage services
Requirements
- Formal training or certification on infrastructure engineering concepts and 3+ years applied experience
- Demonstrated experience using enterprise-authorized AI capabilities within the work environment to support infrastructure engineering workflows, with strong validation habits and awareness of data sensitivity
- Ability to review and validate AI-assisted recommendations before implementation, escalating when uncertain and ensuring outcomes align to resiliency, security, and auditability expectations
- Strong knowledge of storage fundamentals including RAID and erasure coding, replication, snapshots, tiering and caching, IOPS and latency, multipathing, SAN and NAS, and object storage semantics
- Hands-on experience with at least one major storage ecosystem such as NetApp, Dell EMC PowerStore or Isilon, Pure Storage, Hitachi, Ceph, IBM, or cloud-native storage services
- Strong scripting or programming proficiency in one or more of Python, Go, or Bash
- Solid Linux fundamentals including system performance, networking basics, and kernel and storage-stack concepts
- Experience with observability stacks such as Prometheus and Grafana, Elastic or OpenSearch, Splunk, Datadog, or OpenTelemetry
- Proven incident management skills and ability to operate effectively within an on-call rotation
- Practical skills in AI and data operations including anomaly detection, forecasting, correlation, classification, feature extraction, and integrating AI into production tooling and continuous integration and delivery pipelines with safe large language model use, guardrails, and human-in-the-loop review
Preferred qualifications, capabilities, and skills
- Experience with Kubernetes storage using the Container Storage Interface, stateful workloads, and container platform operations
- Proficiency with Infrastructure as Code tools such as Terraform or CloudFormation, and configuration management tools such as Ansible, Chef, or Puppet
- Familiarity with streaming and queue tooling for telemetry and event pipelines such as Kafka
- Experience with IT service management and event management platforms such as ServiceNow
- Knowledge of backup and disaster recovery products and strategy design, including recovery point objective and recovery time objective tradeoffs
- Experience with security controls for data platforms including key management services, hardware security modules, secrets management, and key rotation
Benefits & conditions
We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process.
About the company
You belong to the top echelon of talent in your field. At JPMorganChase, infrastructure isn’t just a foundation - it’s a competitive advantage. This is your opportunity to bring deep storage expertise to a team that operates at global scale, where your contributions directly impact the stability and performance of critical financial services., JPMorganChase, one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world’s most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management., Our professionals in our Corporate Functions cover a diverse range of areas from finance and risk to human resources and marketing. Our corporate teams are an essential part of our company, ensuring that we’re setting our businesses, clients, customers and employees up for success.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
How to Become an AI Engineer
Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production
7 Cloud Computing Trends Coming in 2025 for Developers
What Industries Outside of AI Are Hiring The Most AI Experts?