> Markdown version of [/jobs/ext/2165737-manager-cloud-services-and-site-reliability](https://www.wearedevelopers.com/jobs/ext/2165737-manager-cloud-services-and-site-reliability). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Manager, Cloud Services and Site Reliability - **Company:** Barracuda Networks, Inc. - **Location:** Ann Arbor, MI, United States (Remote available) - **Experience:** Expert - **Salary:** $151,000.0 - $200,000.0 - **Contract:** Permanent contract - **Skills:** Application Portfolio Management, Software as a Service, Cloud Computing, Continuous Integration, DevOps, Disaster Recovery, Distributed Systems, Operational Data Store, Reliability Engineering, Cloud Services, Multi-Cloud, Infrastructure Automation Frameworks - **Published:** August 21, 2026 - **Apply:** https://www.careerjet.com/job/us7f4d308bc892c1f126f8ea411fadabb6/eaa ## About the Role 5+ years of experience in SRE, DevOps, infrastructure, cloud operations, or a related technical operations discipline, including experience leading or managing technical teams. Strong understanding of cloud platforms, distributed systems, production operations, and modern reliability practices. Experience implementing or improving SLOs, SLIs, monitoring, alerting, incident response, and post-incident review practices. Demonstrated ability to hire, mentor, coach, and develop engineers while building a healthy, accountable, and inclusive team culture. Strong communication skills, with the ability to explain technical topics clearly to engineering partners, product stakeholders, and business leaders. Track record of using data, operational insight, and structured problem solving to improve service reliability and team effectiveness. Experience with infrastructure automation, CI/CD practices, disaster recovery, cost optimisation, or multi-cloud operations. Experience influencing operational change across teams, improving documentation practices, or evaluating tools and vendors that support service reliability. ## Description We are looking for a Manager, Site Reliability Engineering to lead a team responsible for the reliability, availability, scalability, and operational excellence of high-volume, business-critical SaaS applications. This role combines people leadership with strong technical judgment, helping the team improve reliability practices, reduce operational toil, and support resilient customer-facing services. In this role, you will manage and develop SRE talent, partner closely with engineering, product, platform, and security teams, and help drive measurable improvements in service health, incident response, automation, and operational readiness. The application portfolio includes products such as Email Security Gateway and Cloud Email Archiving. What you will be working on Lead, coach, and develop a high-performing SRE team, setting clear expectations, supporting career growth, and fostering a culture of ownership, collaboration, and continuous improvement. Drive reliability engineering practices across critical services, including SLOs, SLIs, monitoring, alerting, capacity planning, and service health reporting. Partner with engineering and platform teams to improve the design, operation, and scalability of cloud-based systems, with a focus on reliability, resilience, and maintainability. Own and improve incident management practices, including major incident coordination, post-incident reviews, follow-up actions, and systemic reliability improvements. Champion automation and tooling that reduce manual effort, improve operational consistency, and help the team scale support for production services. Use operational data, service metrics, and risk indicators to identify reliability gaps, prioritize improvements, and communicate progress to technical and business stakeholders. Support secure and compliant operations by partnering with security and engineering teams to embed appropriate controls, documentation, and operational practices into service delivery., Job Description The Food Service Worker will assist the manager with food/meal preparation; maintain cash receipts and meal records. Assist manager in completing daily reports. M… + 2 days ago, Job Description The Food Service Worker will assist the manager with food/meal preparation; maintain cash receipts and meal records. Assist manager in completing daily reports. M… + 16 days ago ## Related Videos - [Leverage Cloud Computing Benefits with Serverless Multi-Cloud ML ](https://www.wearedevelopers.com/videos/78-leverage-cloud-computing-benefits-with-serverless-multi-cloud-ml) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [The Open-source Java SDK for Multi-Cloud Development - Sandeep Pal](https://www.wearedevelopers.com/videos/2113-the-open-source-java-sdk-for-multi-cloud-development-sandeep-pal) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [We adopted DevOps and are Cloud-native, Now What?](https://www.wearedevelopers.com/videos/485-we-adopted-devops-and-are-cloud-native-now-what) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) ## Related Articles - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Events like RSAC Get You CISOs. Developers Decide What Actually Gets Deployed.](https://www.wearedevelopers.com/magazine/693-events-like-rsac-get-you-cisos-developers-decide-what-actually-gets-deployed) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)