> Markdown version of [/jobs/ext/2097443-lead-site-reliability-engineering-network](https://www.wearedevelopers.com/jobs/ext/2097443-lead-site-reliability-engineering-network). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Site Reliability Engineering - Network - **Company:** JPMorgan Chase & Co. - **Location:** Columbus, OH, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Proxy Servers, Microsoft Azure, Border Gateway Protocol, Cloud Computing, Complex Networks, Continuous Integration, Data Centers, DevOps, Failover, Failure Mode Effects Analysis, Monitoring of Systems, Network Security, Network Protocols, Reliability Engineering, Broadcom, Prometheus, Data Driven Tests, Service Discovery, TCP/IP, Wide Area Networks, Load Balancing, Computer Network Technologies, System Availability, Grafana, Reliability of Systems, Firewalls (Computer Science), Juniper, Gitlab, Network Support, Kibana, Terraform, Splunk, Cisco, Jenkins - **Published:** August 17, 2026 - **Apply:** https://www.adzuna.com/details/5830369170 ## About the Role * Formal training or certification in network engineering concepts and 5+ years of applied experience. * 10+ years of experience leading technologists to manage and solve complex technical items within your domain of expertise. * Advanced proficiency in network reliability engineering, including Permit to Operate, FMEA, and operational readiness processes. * Experience leading technologists to manage and solve complex network issues at a firmwide level. * Ability to influence team culture by championing innovation and change for success. * Proficiency in SD-WAN, cloud platforms (AWS, Azure, etc.), and major network technologies (Palo Alto, Juniper, F5, Broadcom, Arista, Cisco, etc.). * Proficiency in observability and monitoring tools such as Grafana, SevOne, Prometheus, Kibana, ThousandEyes, and Splunk. Preferred qualifications, capabilities, and skills * CCIE, Load-balancing, SD-WAN, Observability tools, eBPF, Cloud certs * Demonstrated proficiency in troubleshooting and supporting complex networking environments, including Tier-3 operational support for major incidents. * Experience with continuous integration and delivery tools (e.g., Jenkins, GitLab, Terraform, etc.). * Experience in scalable networking design, including high availability, redundancy, failover, and load balancing. * Experience troubleshooting networking protocols such as TCP/IP, HTTPS, and BGP. * Experience in customer-facing migration, including service discovery, assessment, planning, execution, and operations. This position is subject to Section 19 of the Federal Deposit Insurance Act. As such, an employment offer for this position is contingent on JPMorganChase's review of criminal conviction history, including pretrial diversions or program entries. ## Description Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Network Product , you hold a leadership role in your team, demonstrate strong knowledge across multiple technical domains, and advise others on the technical and business issues facing them. Take lead and conduct resiliency design reviews, break up complex problems into digestible work for other engineers, act as a technical lead for medium to large-sized products, and provide advice and mentoring to other engineers., * Applies network reliability principles (Permit to Operate, FMEA, operational readiness), balancing feature delivery, efficiency, and stability. * Partners with network engineering domains (Datacenter, Firewall, Proxies, DMZ, Load Balancing, etc.) and Lines of Business to align goals and outcomes. * Drives adoption of reliability best practices and observability, demonstrating impact through stability/reliability metrics. * Bridges Engineering, Operations, DevOps, and customers to build resilient, scalable, and secure network services. * Provides Tier-3 network support, leading major incident response, rapid restoration, RCA, and follow-through on corrective actions. * Leads reliability and stability initiatives using data-driven analysis to improve service levels and reduce recurring failure modes. * Defines SLI/SLOs and error budgets with stakeholders and customers, ensuring measurable performance targets and trade-off clarity. * Identifies and removes technical bottlenecks within core domains of expertise, proactively preventing reliability and capacity risks. * Runs blameless, data-driven post-mortems and debriefs, converting learnings (successes and failures) into actionable improvements. * Fosters continuous improvement and strong knowledge sharing, soliciting real-time feedback, avoiding duplicated work, and promoting innovation via internal communities. * Produces and packages thought leadership with specialists/product/engineering teams - documenting best practices and lessons learned for internal assets and industry forums/conferences. ## Related Videos - [How Cisco embraced a DevOps culture within its network engineering team](https://www.wearedevelopers.com/videos/99-how-cisco-embraced-a-devops-culture-within-its-network-engineering-team) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [Enterprise-Cloud-Native - Fast-Paced Development & Deployment in a Highly Secure Banking Environment](https://www.wearedevelopers.com/videos/671-enterprise-cloud-native-fast-paced-development-deployment-in-a-highly-secure-banking-environment) - [Turning Container security up to 11 with Capabilities](https://www.wearedevelopers.com/videos/718-turning-container-security-up-to-11-with-capabilities) - [Inside Bitpanda's Tech Stack: Scaling a European Fintech Leader - Markus Dorner](https://www.wearedevelopers.com/videos/1979-inside-bitpanda-s-tech-stack-scaling-a-european-fintech-leader-markus-dorner) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Best Companies to work for in London: Top 25 Companies in 2023](https://www.wearedevelopers.com/magazine/187-best-companies-to-work-for-in-london-top-25-companies-in-2023) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Where to Find Entry-Level Software Engineering Jobs](https://www.wearedevelopers.com/magazine/397-where-to-find-entry-level-software-engineering-jobs) - [How Much FAANG Companies Actually Pay Software Engineers in 2025](https://www.wearedevelopers.com/magazine/230-how-much-faang-companies-actually-pay-software-engineers-in-2025)