> Markdown version of [/jobs/ext/1704362-manager-site-reliability-engineering](https://www.wearedevelopers.com/jobs/ext/1704362-manager-site-reliability-engineering). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Manager, Site Reliability Engineering - **Company:** Litera Corporation - **Location:** Denver, CO, United States - **Experience:** Experienced - **Salary:** $120,000.0 - $160,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Software as a Service, Configuration Management, Continuous Integration, Disaster Recovery, Performance Tuning, Reliability Engineering, Ansible, Infrastructure Automation Frameworks, Information Technology, Hardware Infrastructure, Terraform - **Published:** July 9, 2026 - **Apply:** https://www.dice.com/job-detail/1ebd9880-de55-4aae-8037-3371fb5494cb ## About the Role * 6+ years of experience operating and troubleshooting distributed SaaS applications across hybrid environments, including on-premises infrastructure and Azure (preferred) or AWS cloud platforms. * 2+ years of SRE management experience leading teams that support multi-product platforms, including experience with resource planning, workload prioritization, team development, and scaling teams. * Strong leadership experience managing major incidents, serving as an escalation point, leading blameless RCAs, and driving continuous post-incident improvements. * Deep experience with observability, monitoring, alerting, APM tools, and reliability metrics, including SLOs, SLIs, and platform health reporting. * Strong technical background in infrastructure automation, configuration management, CI/CD practices, and tools such as Terraform, Ansible, or similar technologies. * Excellent communication, prioritization, and cross-functional collaboration skills, with the ability to align stakeholders and guide teams through complex operational challenges. Nice to Haves: * Previous experience in a Site Reliability Engineering function within a SaaS organization. * Familiarity with ITIL, Problem Management, Change Management, Operational Excellence, or modern platform engineering practices. * Experience supporting highly available customer-facing platforms in a regulated, security-focused, or compliance-driven environment. * Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field; equivalent practical experience; or AWS/Azure professional-level cloud certifications. ## Description This is a hybrid role based in Denver, CO with the expectations to be in office at least 3 days a week for collaboration and connection. Why this Role Matters The Manager, Site Reliability Engineering plays a critical role in ensuring Litera's platforms remain reliable, scalable, secure, and high performing for our customers. This role helps reduce operational risk, improve service availability, and strengthen the systems that support Litera's continued growth. By leading a team of SREs and partnering closely across Engineering, Product, Security, and IT, this leader will drive meaningful improvements in platform resilience, incident response, automation, and operational excellence. The impact of this role is directly tied to customer trust, system stability, and Litera's ability to deliver dependable SaaS solutions at scale. What You'll Deliver * Lead and develop a high-performing SRE team, creating a culture of ownership, collaboration, accountability, and continuous improvement. * Improve the reliability, availability, and performance of Litera's cloud and on-premises platforms through strong operational leadership and technical oversight. * Drive effective incident response practices, including major incident leadership, escalation management, root cause analysis, and post-incident improvement. * Establish and advance reliability metrics, including SLOs, SLIs, platform health indicators, dashboards, and reporting that improve visibility and decision-making. * Reduce operational toil by identifying and implementing automation opportunities across runbooks, monitoring, diagnostics, infrastructure, and support processes. * Partner with Engineering teams to improve application reliability, observability, scalability, and operational readiness across customer-facing systems. * Support platform growth through capacity planning, performance optimization, disaster recovery readiness, and cloud/infrastructure cost optimization. * Collaborate cross-functionally with Product, Engineering, Security, and IT to address operational risks, strengthen service delivery, and align reliability initiatives to business priorities. We're committed to creating an inclusive environment. If you need accommodations at any point in the process or in the role, we're here to support you. ## Related Videos - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Developer Tools for Microsoft Azure](https://www.wearedevelopers.com/videos/450-developer-tools-for-microsoft-azure) - [Terraform for Developers](https://www.wearedevelopers.com/videos/3-terraform-for-developers) - [Implementing Feature Environments with AWS and Terraform](https://www.wearedevelopers.com/videos/531-implementing-feature-environments-with-aws-and-terraform) ## Related Articles - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [The Best Software Developer Blogs to Read](https://www.wearedevelopers.com/magazine/156-the-best-software-developer-blogs-to-read) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers)