> Markdown version of [/jobs/ext/2853683-technology-incident-manager](https://www.wearedevelopers.com/jobs/ext/2853683-technology-incident-manager). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Technology Incident Manager - **Company:** Horizontal Talent - **Location:** Brooklyn, OH, United States - **Experience:** Experienced - **Salary:** $104,000.0 - $116,480.0 - **Contract:** Permanent contract - **Skills:** Distributed Systems, Mainframes - **Published:** September 11, 2026 - **Apply:** https://www.jofdav.com/jobs/59663796-technology-incident-manager ## About the Role * 3+ years of experience leading technical projects, incidents, or cross-functional initiatives * Background supporting highly impactful incidents across multiple business areas or technology teams * Working knowledge of distributed systems, networks, application environments, and mainframe platforms * Familiarity with ITIL-based incident management practices and service restoration workflows * Strong decision-making skills and the confidence to guide response efforts in time-sensitive situations * Excellent communication skills for both technical and non-technical audiences * Ability to stay organized, composed, and effective under pressure * Detail-oriented documentation and follow-through skills * Bachelor's degree in a related business or science field, or equivalent work experience Preferred Skills * ITIL or incident management certification * Project management certification such as PMP, or a related technical certification * Experience facilitating high-stakes incident calls and driving cross-functional resolution * Exposure to production readiness reviews and recovery documentation maintenance * Experience identifying process gaps and contributing to continuous improvement initiatives * Ability to adapt quickly in a shifting environment while maintaining professionalism and collaboration Horizontal is committed to fostering an inclusive, respectful, and equitable workplace where diverse perspectives are valued and all candidates are supported throughout the hiring process. We encourage applicants from all backgrounds to apply and bring their unique experiences to the team. ## Description Join a dynamic incident management team where you'll help restore critical technology services, coordinate cross-functional response efforts, and keep stakeholders informed during high-impact events. This role is ideal for someone who thrives in fast-paced environments, enjoys solving complex problems, and is confident guiding recovery efforts across technical teams. Responsibilities * Monitor, assess, and help drive resolution of critical technology incidents and outages affecting enterprise operations * Proactively investigate potential incidents and engage the right technical support partners as issues emerge * Lead incident response activities, including triage calls, bridge coordination, and vendor collaboration * Communicate clearly and consistently with business stakeholders, technical teams, and leadership throughout the incident lifecycle * Escalate urgent issues appropriately and provide timely progress updates during troubleshooting and remediation * Document incident timelines, restoration actions, and follow-up items to support post-incident reviews and process improvement * Partner with teams across monitoring, detection, problem management, and communications to strengthen recovery readiness * Support production readiness efforts for critical initiatives and help improve incident management processes and metrics * Participate in rotating on-call support and off-hours coverage, including overnight and weekend needs ## Related Videos - [Applying Agile Principles to Incident Management ](https://www.wearedevelopers.com/videos/101-applying-agile-principles-to-incident-management) - [Cloud as the new mainframe: why the cloud hype does not reflect the dev reality](https://www.wearedevelopers.com/videos/797-cloud-as-the-new-mainframe-why-the-cloud-hype-does-not-reflect-the-dev-reality) - [Best Practices for AI-Assisted Development of Distributed Systems](https://www.wearedevelopers.com/videos/100200-best-practices-for-ai-assisted-development-of-distributed-systems) - [What makes Cybersecurity different for critical infrastructure?](https://www.wearedevelopers.com/videos/571-what-makes-cybersecurity-different-for-critical-infrastructure) - [".Net is too slow" was not an option](https://www.wearedevelopers.com/videos/100176-net-is-too-slow-was-not-an-option) - [Your Distributed System Just Got a Brain. Now What?](https://www.wearedevelopers.com/videos/100017-your-distributed-system-just-got-a-brain-now-what) ## Related Articles - [The Geometry of Incidents: Connecting User Impact to Architecture](https://www.wearedevelopers.com/magazine/764-the-geometry-of-incidents-connecting-user-impact-to-architecture) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Should Tech Managers Be Developers First? Pros and Cons](https://www.wearedevelopers.com/magazine/327-should-tech-managers-be-developers-first-pros-and-cons) - [From developer to manager – what does it take to become an engineering manager?](https://www.wearedevelopers.com/magazine/42-from-developer-to-manager-what-does-it-take-to-become-an-engineering-manager) - [System change: restart as developer?](https://www.wearedevelopers.com/magazine/39-system-change-restart-as-developer) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)