> Markdown version of [/jobs/ext/3577964-site-reliability-engineer-ii](https://www.wearedevelopers.com/jobs/ext/3577964-site-reliability-engineer-ii). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer II - **Company:** Early Warning Services, LLC. - **Location:** Chicago, IL, United States - **Experience:** Experienced - **Salary:** $83,000.0 - $110,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Systems Engineering, Microsoft Azure, Cloud Computing, Information Systems, Computer Engineering, Continuous Integration, Linux, DevOps, Disaster Recovery, Distributed Systems, Reliability Engineering, Software Engineering, Data Logging, Scripting, Google Cloud, Cloud Platform System, Infrastructure Automation Frameworks, Information Technology, Oracle Cloud Infrastructure, Programming Languages - **Published:** October 4, 2026 - **Apply:** https://startup.jobs/site-reliability-engineer-ii-early-warning-services-llc-10272553 ## About the Role Candidates responding to this posting must independently possess the eligibility to work in the United States, for any employer, at the date of hire. This position is ineligible for employment Visa sponsorship., * Typically 2-5 years of relevant professional experience in Software Engineering, Site Reliability Engineering, Systems Engineering, Cloud/Platform Engineering, DevOps, Infrastructure Engineering, Architecture where applicable, or a comparable technical discipline. * Experience with software development or scripting using one or more modern programming languages. * Experience with software engineering principles, distributed systems, production troubleshooting, automation, and observability appropriate to the level. * Experience with public cloud technologies and architectures, preferably AWS, along with infrastructure, networking, Linux/Unix, and modern application architectures appropriate to the level. * Demonstrated analytical, problem-solving, communication, and collaboration skills appropriate to the scope of the role., * Hands-on experience with AWS is preferred, or comparable experience with another major cloud platform such as Microsoft Azure, Google Cloud Platform (GCP), or Oracle Cloud Infrastructure (OCI). * Experience developing, deploying, operating, or improving highly available production software or distributed systems. * Experience with CI/CD, Infrastructure as Code, containers or orchestration, observability, monitoring, alerting, and software-delivery automation. * Experience with SLIs, SLOs, error budgets, incident management, performance analysis, capacity management, resilience testing, disaster recovery, or operational readiness appropriate to the level. * Experience creating reusable automation, tooling, platforms, patterns, or practices that improve engineering effectiveness. * Bachelor's degree in Computer Science, Software Engineering, Computer Engineering, Information Systems, or a related technical field, or equivalent practical experience. Physical Requirements Working conditions consist of a normal office environment. Work is primarily sedentary and requires extensive use of a computer and involves sitting for periods of approximately four hours. Work may require occasional standing, walking, kneeling and reaching. Must be able to lift 50 pounds occasionally and/or negligible amount of force frequently. Requires visual acuity and dexterity to view, prepare, and manipulate documents and office equipment including personal computers. Requires the ability to communicate with internal and/or external customers. Employee must be able to perform essential functions and physical requirements of position with or without reasonable accommodation. ## Description JobPosting MonetaryAmount USD QuantitativeValue 132000 83000 YEAR 2026-10-02T00:00:00Z At Early Warning, we've powered and protected the U.S. financial system for over thirty years with cutting-edge solutions like Zelle®, Paze®, and so much more. As a trusted name in payments, we partner with thousands of institutions to increase access to financial services and protect transactions for hundreds of millions of consumers and small businesses. Early Warning follows a hybrid work model to allow for a more collaborative working environment., * The Site Reliability Engineer II applies software engineering and systems engineering practices to improve the reliability, resilience, scalability, and operational health of production services. The role partners with Software Engineering and other technology teams to ensure reliability, observability, recoverability, performance, and operational readiness are engineered into systems throughout their lifecycle. At the Engineer II level, the engineer works with increasing independence, identifies reliability risks, automates operational work, and helps guide reliability decisions within assigned services. Core Responsibilities * Use software engineering, automation, and DevOps principles and practices to continually improve how services are built, tested, deployed, observed, operated, and recovered. * Use data, evidence, experimentation, and rigorous engineering analysis appropriate to the level to identify reliability risks, test assumptions, and guide technical decisions. * Define, implement, or improve SLIs, SLOs, error budgets, and other service-health measures appropriate to the scope of responsibility. * Improve observability through metrics, logging, tracing, monitoring, alerting, dashboards, and service-health instrumentation. * Drive continuous improvement across CI/CD, observability, deployment practices, Infrastructure as Code, automation, testing, incident response, capacity management, resilience, and operational readiness. * Identify recurring or systemic production issues and translate operational experience into improvements in code, architecture, automation, tooling, and engineering practices. * Partner with Software Engineering teams to incorporate reliability, resiliency, scalability, performance, observability, recoverability, and operational readiness throughout the development lifecycle. * Participate in or lead incident response, troubleshooting, service restoration, and blameless post-incident learning appropriate to the level. * Participates independently in an on-call rotation for supported production services, diagnosing and resolving incidents and service degradation and escalating complex issues as appropriate. * Reduce operational toil and unnecessary manual intervention through software, automation, reusable patterns, and better engineering practices. Leveling Intent Engineer II performs the same fundamental SRE mission as Engineer I with greater independence, proficiency, and responsibility for increasingly complex systems. The engineer increasingly multiplies team effectiveness through reusable solutions, knowledge transfer, and improvements that reduce repeated work. Level Expectations * Expands their force-multiplier impact by creating reusable solutions, sharing expertise, and helping teammates avoid repeatedly solving the same operational and reliability problems. * Demonstrates software engineering, systems thinking, troubleshooting, and production reliability capabilities appropriate to the level. * Applies evidence-driven reasoning and technical rigor to distinguish observed facts from assumptions and make defensible engineering recommendations. * Shares knowledge and contributes to sustainable engineering capability rather than creating dependency on individual expertise. * Works independently on well-defined reliability problems and increasingly on more complex issues. * Identifies and implements continuous improvements within assigned services and areas of responsibility. ## Related Videos - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Shifting Stress to Progress— Understanding DevOps to do DevOps Better](https://www.wearedevelopers.com/videos/268-shifting-stress-to-progress-understanding-devops-to-do-devops-better) - [JavaScript? No. Java Scripts! - Scripting with Java](https://www.wearedevelopers.com/videos/2094-javascript-no-java-scripts-scripting-with-java) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025) - [Where to Find Entry-Level Software Engineering Jobs](https://www.wearedevelopers.com/magazine/397-where-to-find-entry-level-software-engineering-jobs)