> Markdown version of [/jobs/ext/3002550-lead-software-engineer](https://www.wearedevelopers.com/jobs/ext/3002550-lead-software-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Software Engineer - **Company:** State Farm Insurance - **Location:** Bloomington, IL, United States (Remote available) - **Experience:** Expert - **Salary:** $120,000.0 - $160,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Amazon Web Services, Automation of Tests, Continuous Integration, Distributed Systems, Fault Tolerance, Python (Programming Language), OpenShift, Reliability Engineering, Software Engineering, Software Vulnerability Management, Performance Testing, Spring Cloud, Infrastructure as Code (IaC), Git Flow, Kubernetes, Deployment Automation, Terraform, Dynatrace, Programming Languages - **Published:** September 19, 2026 - **Apply:** https://www.techcareers.com/job.asp?id=3397305302&tx=YT1212THT&pt=1&aff=0B19D771-A501-4A5E-8338-2A822B784D54&utm_source=Job%20Feed&utm_medium=textkernel&utm_campaign=DE&utm_term=0B19D771-A501-4A5E-8338-2A822B784D54 ## About the Role * Significant hands-on software engineering experience designing, developing, testing, deploying, troubleshooting, and supporting production applications using modern programming languages such as Java, Python, or similar languages. * Experience designing and operating resilient distributed systems, including practical application of patterns for fault tolerance, dependency failures, scalability, recovery, and graceful degradation. * Hands-on AWS experience developing and operating cloud-native applications and services. * Experience with containers and orchestration platforms such as Kubernetes or OpenShift; Red Hat OpenShift Service on AWS (ROSA) experience is highly preferred. * Strong working knowledge of Site Reliability Engineering (SRE) principles and demonstrated experience applying resiliency, observability, operational readiness, and incident-management practices to production applications. * Experience with observability and distributed tracing, including using telemetry to troubleshoot complex application and dependency issues and establishing effective monitoring and alerting; Dynatrace experience is preferred, with comparable experience using similar platforms also considered. * Experience with modern software delivery practices including CI/CD, automated testing, Git-based workflows, deployment automation, security/vulnerability management, and Infrastructure as Code (IaC) using OpenTofu, Terraform, or similar technologies. * Experience conducting root-cause analysis, resolving complex production issues, and translating incident findings into sustainable engineering improvements. * Experience using AI-assisted engineering capabilities to improve software development, testing, analysis, troubleshooting, automation, or operational workflows. * Demonstrated technical leadership, communication, mentoring, and consulting skills, including the ability to assess technical risk and influence engineering teams without direct authority. ## Description The Digital Experience (DE) Resiliency & Availability team is seeking two experienced Lead Software Engineers to improve the reliability, resiliency, and engineering effectiveness of critical customer-facing digital experiences., Both roles are hands-on technical leadership positions that work across engineering teams to solve complex problems and improve how applications are built and operated. The positions have two complementary focus areas: Engineering Excellence, focused on helping teams improve engineering practices and operational maturity; and Reliability Improvement, focused on using engineering telemetry and hands-on problem solving to improve customer-visible availability and performance. Candidates may be considered for either role based on their experience and interests. * Partner with engineering teams to design, build, troubleshoot, and improve resilient, highly available production systems. * Provide hands-on technical leadership and consulting across product teams, using engineering expertise to diagnose complex problems and guide implementation of sustainable solutions. * Apply Site Reliability Engineering (SRE) principles to improve availability, performance, observability, incident readiness, and operational maturity. * Use observability and distributed tracing to identify failure patterns, troubleshoot dependencies, measure customer impact, and drive data-informed reliability improvements. * Assess engineering practices and coach teams on automated testing, deployment readiness, monitoring and alerting, production support, and other engineering best practices. * Participate in production incident analysis and post-incident reviews, identifying root causes and driving corrective actions that prevent recurrence. * Build tooling, automation, dashboards, and AI-assisted engineering capabilities that identify risk earlier and improve engineering effectiveness. * Support resiliency exercises, performance testing, GameDays, chaos testing, and other practices that validate how applications behave under failure or degraded conditions. * Collaborate with engineers, architects, SRE and platform teams, and technology leaders across organizational boundaries while mentoring engineers and influencing technical direction. ## Related Videos - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [The Power of Purpose: Unlocking Potential and Innovation](https://www.wearedevelopers.com/videos/1110-the-power-of-purpose-unlocking-potential-and-innovation) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [Next Level Enterprise Architecture: Modular, Flexible, Scalable, Multichannel and AI-Ready?](https://www.wearedevelopers.com/videos/1017-next-level-enterprise-architecture-modular-flexible-scalable-multichannel-and-ai-ready) - [It's Not Vibe Coding If You Know What You're Doing](https://www.wearedevelopers.com/videos/100119-it-s-not-vibe-coding-if-you-know-what-you-re-doing) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Best Countries for Software Engineers](https://www.wearedevelopers.com/magazine/267-best-countries-for-software-engineers) - [How Much FAANG Companies Actually Pay Software Engineers in 2025](https://www.wearedevelopers.com/magazine/230-how-much-faang-companies-actually-pay-software-engineers-in-2025) - [What is Software Engineering?](https://www.wearedevelopers.com/magazine/289-what-is-software-engineering)