SRE Software Engineer III

JPMorgan Chase & Co.
Jersey City, NJ, United States
18 days ago
Apply on jpmc.fa.oraclecloud.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Java (Programming Language) .NET Framework Artificial Intelligence Automation of Tests Unit Testing Code Generation Software Quality Computer Networks Continuous Delivery Continuous Integration Distributed Systems Fault Tolerance
+23 more
Python (Programming Language) Reliability Engineering Software Tools Prometheus Runbook Secure Coding Software Engineering Software Systems Toolchain Datadog Delivery Pipeline Grafana Spring-boot AWS ECS Rate Limiting Gitlab Kubernetes Terraform Splunk Dynatrace Docker Jenkins Programming Languages

Job description

  • Design and implement automated continuous integration and continuous delivery pipelines to improve release quality, speed, and repeatability
  • Develop, test, and deliver software solutions that strengthen availability, reliability, scalability, and operational readiness of applications and services
  • Partner with engineers, technical experts, and key stakeholders to troubleshoot and resolve complex, multi-system problems to restore service and prevent recurrence
  • Define and use service level indicators and service level objectives to proactively identify risk, prioritize reliability improvements, and reduce customer impact
  • Advance observability practices by improving telemetry, dashboards, and alerting to shorten time-to-detect and time-to-recover
  • Improve operational excellence through runbooks, automation, and continuous improvement actions that reduce toil and production risk
  • Leverages enterprise-authorized AI coding assist tools within the work environment to improve code quality, delivery speed, and productivity across complex deliverables (e.g., code generation/refactoring, unit test creation, documentation), while validating outputs through peer review, automated testing, and secure coding standards; contributes learnings and reusable patterns to improve broader team effectiveness.
  • Applies knowledge of tools within the Software Development Life Cycle toolchain, including enterprise-authorized AI-assisted development and automation capabilities, to improve the value realized by automation.

Requirements

  • Formal training or certification on software engineering concepts and 3+ years applied experience
  • Proficiency in site reliability culture and principles, including applying reliability practices within an application or platform
  • Proficiency in at least one programming language such as Python, Java/Spring Boot, or .NET
  • Experience building observability solutions, including telemetry collection and service level objective-based alerting using tools such as Grafana, Dynatrace, Prometheus, Datadog, or Splunk
  • Experience with continuous integration and continuous delivery tools such as Jenkins, GitLab, or Terraform
  • Familiarity with containers and orchestration technologies such as Docker, Kubernetes, or Amazon Elastic Container Service
  • Working knowledge of diagnosing and troubleshooting common networking concepts and issues in distributed systems
  • Ability to proactively remove blockers, learn new technologies, and apply new approaches to improve delivery outcomes
  • Working knowledge of using enterprise-authorized AI-assisted software development tools within the work environment (e.g., for coding, test creation, troubleshooting, or documentation) with demonstrated ability to critically evaluate, validate, and refine AI-generated outputs for correctness, performance, and security.
  • Understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; ability to guide peers on safe and effective usage within team practices.

Preferred Qualifications, Capabilities, and Skills

  • Experience designing reliability improvements using error budgets, capacity planning, and resilience patterns (e.g., rate limiting, backpressure, graceful degradation)
  • Experience improving release safety with progressive delivery practices (e.g., canary deployments, feature flags, automated rollbacks)
  • Experience implementing automated quality gates (unit, integration, performance, and security testing) within delivery pipelines
  • Exposure to incident management practices, post-incident reviews, and implementing measurable remediation actions to prevent recurrence
  • Familiarity with infrastructure-as-code and standardized environment provisioning to improve consistency and auditability

Benefits & conditions

We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process.

About the company

JPMorganChase, one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world’s most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jpmc.fa.oraclecloud.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

6:14 min

Structuring CI/CD pipelines with integrated security and quality checks

Christoph Ruggenthaler · LIVE

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all