Manager, Cloud Services and Site Reliability

Barracuda Networks, Inc.
Ann Arbor, MI, United States
16 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$151,000.0 - $200,000.0
Working hours
Regular working hours

Tech stack

Application Portfolio Management Software as a Service Cloud Computing Continuous Integration DevOps Disaster Recovery Distributed Systems Operational Data Store Reliability Engineering Cloud Services Multi-Cloud Infrastructure Automation Frameworks

Job description

We are looking for a Manager, Site Reliability Engineering to lead a team responsible for the reliability, availability, scalability, and operational excellence of high-volume, business-critical SaaS applications. This role combines people leadership with strong technical judgment, helping the team improve reliability practices, reduce operational toil, and support resilient customer-facing services. In this role, you will manage and develop SRE talent, partner closely with engineering, product, platform, and security teams, and help drive measurable improvements in service health, incident response, automation, and operational readiness. The application portfolio includes products such as Email Security Gateway and Cloud Email Archiving. What you will be working on Lead, coach, and develop a high-performing SRE team, setting clear expectations, supporting career growth, and fostering a culture of ownership, collaboration, and continuous improvement. Drive reliability engineering practices across critical services, including SLOs, SLIs, monitoring, alerting, capacity planning, and service health reporting. Partner with engineering and platform teams to improve the design, operation, and scalability of cloud-based systems, with a focus on reliability, resilience, and maintainability. Own and improve incident management practices, including major incident coordination, post-incident reviews, follow-up actions, and systemic reliability improvements. Champion automation and tooling that reduce manual effort, improve operational consistency, and help the team scale support for production services. Use operational data, service metrics, and risk indicators to identify reliability gaps, prioritize improvements, and communicate progress to technical and business stakeholders. Support secure and compliant operations by partnering with security and engineering teams to embed appropriate controls, documentation, and operational practices into service delivery., Job Description The Food Service Worker will assist the manager with food/meal preparation; maintain cash receipts and meal records. Assist manager in completing daily reports. M…

  • 2 days ago, Job Description The Food Service Worker will assist the manager with food/meal preparation; maintain cash receipts and meal records. Assist manager in completing daily reports. M…
  • 16 days ago

Requirements

5+ years of experience in SRE, DevOps, infrastructure, cloud operations, or a related technical operations discipline, including experience leading or managing technical teams. Strong understanding of cloud platforms, distributed systems, production operations, and modern reliability practices. Experience implementing or improving SLOs, SLIs, monitoring, alerting, incident response, and post-incident review practices. Demonstrated ability to hire, mentor, coach, and develop engineers while building a healthy, accountable, and inclusive team culture. Strong communication skills, with the ability to explain technical topics clearly to engineering partners, product stakeholders, and business leaders. Track record of using data, operational insight, and structured problem solving to improve service reliability and team effectiveness. Experience with infrastructure automation, CI/CD practices, disaster recovery, cost optimisation, or multi-cloud operations. Experience influencing operational change across teams, improving documentation practices, or evaluating tools and vendors that support service reliability.

Benefits & conditions

A team where you can voice your opinion, make an impact, and where you and your experience are valued. Internal mobility - there are opportunities for cross training and the ability to attain your next career step within Barracuda. Equity, in the form of non-qualifying options High-quality health benefits Retirement Plan with employer match Career-growth opportunities Flexible Time Off and Paid Time Off benefits Volunteer opportunities The anticipated salary range for this role is $151,000 to $200,000. Actual compensation offered will be dependent upon the individual’s skills, experience, and qualifications as they directly relate to the requirements of the position, the budget for the position, and applicable employment laws. At Barracuda, we believe in fair and equitable compensation practices that reflect both market realities and the unique circumstances of each geographical location. We recognize that cost-of-living disparities, market conditions, and other factors can significantly impact compensation expectations in different regions. The compensation range provided in this job description is for illustrative purposes only and may not reflect the actual compensation offers for the position in your location. Final compensation will be determined based on a variety of factors including the candidates’ qualifications and experience. #LI-Remote, + $20.00-30.00 per hour

About the company

Come join our passionate team! Barracuda is a leading cybersecurity company providing complete protection against complex threats. Our platform protects email, data, applications, and networks with innovative solutions, and a managed XDR service, to strengthen cyber resilience. Hundreds of thousands of IT professionals and managed service providers worldwide trust us to protect and support them with solutions that are easy to buy, deploy, and use. We know a diverse workforce adds to our collective value and strength as an organization. Barracuda Networks is proud to be an Equal Opportunity Employer, committed to equal employment opportunity and equitable compensation regardless of race, gender, religion, sex, sexual orientation, national origin, or disability. Envision yourself at Barracuda

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

1:22 min

Overcoming developer challenges in multi-cloud environments

Sandeep Pal Sandeep Pal · Coffee With Developers

2:41 min

Transitioning artificial intelligence infrastructure into scalable commodity cloud services

juarezjunior juarezjunior · World Congress 2024

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:00 min

Connecting multi-cloud services for automated data preparation

Linda Mohamed · LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all