> Markdown version of [/jobs/ext/1956499-manager-software-engineering-devops](https://www.wearedevelopers.com/jobs/ext/1956499-manager-software-engineering-devops). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Manager, Software Engineering DevOps - **Company:** The Options Clearing Corporation - **Location:** Chicago, IL, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** DevOps, Middleware, Software Engineering, Cloud Platform System, Delivery Pipeline, Mttr, 3-tier Architectures - **Published:** August 6, 2026 - **Apply:** https://us.experteer.com/career/view-jobs/manager-software-engineering-devops-chicago-il-usa-58813774 ## About the Role * for P1 and P2 incidents with identified root causes and action plans * Define, publish, and enforce SLA targets across severity levels and monitor real-time compliance * Publish monthly SLA reports with trend analysis and improvement actions * Lead alert tuning and MTTR initiatives to reduce toil and improve triage efficiency * Identify recurring incident patterns and drive permanent fixes; open and track Problem records * Lead automation and tooling projects to streamline support workflows * Ensure accurate incident documentation and maintain runbooks and knowledge base * Manage a team of 6-10 L1/L2 engineers including on-call coverage and career development * Oversee talent management, performance reviews, and training plans * Collaborate with cross-functional teams to align operational policies and priorities * Maintain up-to-date environment configuration documentation and procedures Tasks * Proven team leadership experience with accountability for standards across incident types and urgencies * Strong cross-functional collaboration skills across L1/L2 Platform Security and App Dev teams * Hands-on experience in production environment supports including deployment pipelines, containers, and middleware * Ability to create, tune, and maintain monitoring alerts and runbooks independently * Excellent oral and written communication skills for leadership reporting * Analytical, judgement, and consultation skills under pressure with stakeholder management * Ability to manage multiple priorities with strong organizational discipline * Minimum 5 years in environment operations or related fields with interdisciplinary exposure Key requirements * hybrid work environment (remote up to 2 days/week) * tuition reimbursement * student loan repayment assistance * technology stipend * generous PTO and parental leave * 401k employer match ## Description Experteer Overview In this role you lead L1/L2 production environment support across deployments, middleware, and platform infrastructure, driving incident response quality and SLA governance. You will deliver actionable metrics and ensure rapid, well-documented triage, resolution, and post-incident learning. You'll manage a team to maintain readiness, escalate complex incidents, and implement automation to reduce toil. This is a high-impact, operations-focused leadership role at OCC with a strong emphasis on reliability and continuous improvement. Compensation / Benefits * Lead L1/L2 incident response activities from triage to closure and post-incident reporting * Oversee technical analysis of environment incidents across deployment, middleware, and platform layers * Serve as Tier 3 escalation point for complex incidents across Platform, S&I, Security, and App Dev teams * Own end-to-end incident lifecycle including RCA and permanent fixes or workarounds * Drive post-incident reviews for P1 and P2 incidents with identified root causes and action plans * Define, publish, and enforce SLA targets across severity levels and monitor real-time compliance * Publish monthly SLA reports with trend analysis and improvement actions * Lead alert tuning and MTTR initiatives to reduce toil and improve triage efficiency * Identify recurring incident patterns and drive permanent fixes; open and track Problem records * Lead automation and tooling projects to streamline support workflows * Ensure accurate incident documentation and maintain runbooks and knowledge base * Manage a team of 6-10 L1/L2 engineers including on-call coverage and career development * Oversee talent management, performance reviews, and training plans * Collaborate with cross-functional teams to align operational policies and priorities * Maintain up-to-date environment configuration documentation and procedures Tasks * Proven team leadership experience with accountability for standards across incident types and urgencies * Strong cross-functional collaboration skills across L1/L2 Platform Security and App Dev teams * Hands-on experience in production environment supports including deployment pipelines, containers, and middleware * Ability to create, tune, and maintain monitoring alerts and runbooks independently * Excellent oral and written communication skills for leadership reporting * Analytical, judgement, and consultation skills under pressure with stakeholder management * Ability to manage multiple priorities with strong organizational discipline * Minimum 5 years in environment operations or related fields with interdisciplinary exposure Key requirements * hybrid work environment (remote up to 2 days/week) * tuition reimbursement * student loan repayment assistance * technology stipend * generous PTO and parental leave * 401k employer match ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Applying Agile Principles to Incident Management ](https://www.wearedevelopers.com/videos/101-applying-agile-principles-to-incident-management) - [What Developers Get Wrong About Application Quality](https://www.wearedevelopers.com/videos/233-what-developers-get-wrong-about-application-quality) - [Developer’s Perspective: Overview of the Tezos Blockchain Ecosystem](https://www.wearedevelopers.com/videos/237-developer-s-perspective-overview-of-the-tezos-blockchain-ecosystem) - [We adopted DevOps and are Cloud-native, Now What?](https://www.wearedevelopers.com/videos/485-we-adopted-devops-and-are-cloud-native-now-what) - [Plan CI/CD on the Enterprise level!](https://www.wearedevelopers.com/videos/544-plan-ci-cd-on-the-enterprise-level) ## Related Articles - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023) - [Now is the time for industrialized software development](https://www.wearedevelopers.com/magazine/601-now-is-the-time-for-industrialized-software-development) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)