> Markdown version of [/jobs/ext/61496-lead-software-engineer-sre](https://www.wearedevelopers.com/jobs/ext/61496-lead-software-engineer-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Software Engineer - SRE - **Company:** Wells Fargo - **Location:** St. Louis, MO, United States - **Experience:** Expert - **Salary:** $119,000.0 - $187,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), JavaScript (Programming Language), Microsoft Windows, Amazon Web Services, Computing Platforms, Microsoft Azure, Software as a Service, Cloud Computing, Linux, Monitoring of Systems, JSON, Python (Programming Language), Node.Js, OpenShift, Performance Tuning, Reliability Engineering, Prometheus, Ruby, Shell Script, Software Engineering, Scripting, System Availability, Grafana, Safety Critical Systems, Reliability of Systems, AngularJS, Kubernetes, Performance Monitor, Heap (Data Structure), Splunk, Appdynamics, Vmware, Programming Languages - **Published:** May 30, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=a5dbb813b991ed4e ## About the Role Do you have experience in Windows support?, * 5+ years of Software Engineering experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education * 5+ years of experience leading observability and monitoring tooling - Splunk, AppDynamics, Splunk Observability, Grafana, Open Telemetry * 5+ years in infrastructure (windows and Linux) support * 5+ years proven success in toil reduction initiatives * 5+ years in cloud application management especially OpenShift Container Platform, * 5+ Years' experience in SRE, public & private cloud technologies, Java performance tuning, capacity optimization for mission critical applications * Working knowledge of multiple programming languages (e.g., Java, JavaScript, Ruby, Python, JSON, Angular, NodeJS) * Hands-on experience with cloud and platform technologies such as AWS, PCF, PKS, Kubernetes, OpenShift, Linux, Azure, Windows, and VMware * Strong verbal, written, and interpersonal communication skills for effective collaboration across teams * Ability to engage with and influence stakeholders at various organizational levels * Expert experience on monitoring tools - Prometheus, Grafana, AppDynamics, Glassbox, Splunk * Advanced experience in one or more scripting languages - Python, Shell scripting etc * Strong knowledge of Kubernetes, OCP and troubleshooting skills * Strong grasp of Java performance concepts (heap, GC) and critical monitoring metrics for Java apps * Ability to identify manual tasks in the processes and automating them to reduce toil Job Expectations: * Willingness to work on-site at stated location on the job opening * This position offers a hybrid work schedule ## Description Wells Fargo is seeking a Lead Site Reliability Engineer (SRE) to join the WIMT Platform team. This role is responsible for driving the stability, resiliency, performance, and security of mission-critical platforms that support Wells Fargo Advisors, First Clearing firms, and FINET practices. As a Lead SRE, you will provide hands-on technical leadership across incident management, automation, observability, and reliability engineering, with a strong focus on proactive risk mitigation and continuous improvement. You will help define and enforce reliability standards while partnering closely with Application Development, Product, Business, and Enterprise teams to ensure operational excellence throughout the full-service lifecycle. This role is ideal for a highly motivated engineer with deep experience operating large-scale, production systems who takes ownership, values accountability, and is passionate about building resilient, enterprise-grade platforms. Learn more about career areas and business divisions at https://www.wellsfargojobs.com. In this role, you will: * Design and implement scalability, reliability, and observability strategies for cloud and on-premise environments * Define SLIs (Service Level Indicators), SLOs (Service Level Objectives), and Error Budgets to improve system reliability * Provide vision, direction and expertise to leadership on implementing innovative and significant business solutions * Maintain knowledge of industry best practices and new technologies and recommend innovations that enhance operations or provide a competitive advantage to the organization * Strategically engage with all levels of professionals and managers across the enterprise and serve as an expert advisor to leadership * Review and analyze complex, large-scale technology solutions for tactical and strategic business objectives, enterprise technological environment, and technical challenges that require in-depth evaluation of multiple factors, including intangibles or unprecedented technical factors * Drive adoption of NFRs, best practices-quality and compliance across observability and performance engineering * Ensure high availability and performance of production systems through proactive monitoring and incident response * Collaborate and consult with key technical experts, senior technology team, and external industry groups to resolve complex technical issues and achieve goals * Lead projects, teams, or serve as a peer mentor ## Related Videos - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [Tips and Tricks for Working with JSON](https://www.wearedevelopers.com/videos/1229-tips-and-tricks-for-working-with-json) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Coroutine explained yet again 60 years later](https://www.wearedevelopers.com/videos/690-coroutine-explained-yet-again-60-years-later) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [How Much Does a Software Engineer Make? Realistic Software Engineering Salaries](https://www.wearedevelopers.com/magazine/425-how-much-does-a-software-engineer-make-realistic-software-engineering-salaries) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Best Countries for Software Engineers](https://www.wearedevelopers.com/magazine/267-best-countries-for-software-engineers) - [How Much FAANG Companies Actually Pay Software Engineers in 2025](https://www.wearedevelopers.com/magazine/230-how-much-faang-companies-actually-pay-software-engineers-in-2025)