> Markdown version of [/jobs/ext/1248012-senior-director-enterprise-reliability](https://www.wearedevelopers.com/jobs/ext/1248012-senior-director-enterprise-reliability). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Director, Enterprise Reliability... - **Company:** Marriott International, Inc. - **Location:** Bethesda, MD, United States (Remote available) - **Experience:** Expert - **Salary:** $151,100.0 - $258,000.0 - **Contract:** Permanent contract - **Skills:** Agile Methodology, Amazon Web Services, Microsoft Azure, Cloud Engineering, Information Systems, DevOps, Distributed Systems, IT Management, Oracle (Applications), Reliability Engineering, Scaled Agile Framework, Software Engineering, Multi-Cloud, Information Technology, Devsecops, Service Stack - **Published:** July 12, 2026 - **Apply:** https://www.juju.com/job/00000000gfkxxe ## About the Role + Undergraduate degree in Engineering, Information Systems or Computer Science discipline and/or equivalent experience/certification + 10+ years of senior IT leadership experience across software engineering, platform engineering, SRE, DevOps, or related disciplines with a blend of deep technical knowledge and a customer-focused mindset + 8+ years leading large-scale engineering or reliability organizations + 3-5 years' experience operating and maintaining a Multi-Cloud environment + 8+ years managing direct reports and Service Provider teams + Demonstrated success driving reliability transformation across multiple engineering domains + Deep understanding of cloud-native architectures, distributed systems, and modern software delivery + Proven ability to influence technical strategy and standards without direct authority + Experience defining SLOs, SLIs, error budgets, and operational maturity frameworks + Strong executive communication skills and systems thinking mindset Preferred: + Graduate Degree in a technical discipline + 3+ years experience managing Public Cloud technology stacks such as AWS, Azure, Alibaba, GCP, etc. + 10+ years experience leading and engaging highly technical architecture, engineering and operations teams + 10+ years of hands on technical experience in infrastructure and/or development teams + Demonstrated ability to manage a large, diverse portfolio of technologies and projects, while balancing short-term goals and long-term vision + Comfortable applying a combination of qualitative and quantitative methods to define success + Expert problem solver that is to be able to solve challenges across people, process and technology + Ability to build strong relationships and network throughout the broader IT community + Experience in the security, implementation and operational support of mission critical products + Strong influencing skills and an ability to overcome barriers while driving change + Experience in researching emerging technologies and trends, standards, and products + Experience in developing technology roadmaps and strategies + Excellent verbal and written communication skills for a wide range of audiences including executives, business stakeholders and IT teams + Strong knowledge of emerging tools, software, applications, and systems for attaining best-in-class IT technology across the enterprise + Strong attention to detail with an ability to operate effectively across multiple priorities + ITILv4 certification + Experience operating in an Agile or SAFE framework + AWS, Azure, Oracle or similar cloud certifications + Familiarity with hospitality, travel, retail, or large-scale consumer-facing platforms CORE WORK ACTIVITIES Domain Reliability Strategy & Transformation + Define and execute domain reliability strategies for Data, and Core platforms + Translate enterprise reliability transformation objectives into actionable domain roadmaps + Establish service ownership models, service tiering, and reliability targets + Identify organizational, process, and technology barriers to reliability adoption and drive remediation SRE & DevSecOps Enablement + Expand and mature embedded SRE capabilities within domain engineering teams + Drive adoption of SLOs, error budgets, production readiness, and operational excellence practices + Embed reliability considerations into architecture reviews and CI/CD pipelines + Improve deployment safety, change quality, and risk reduction Incident Reduction, Resiliency & Learning ## Description The Senior Director, Enterprise Reliability Engineering is responsible for improving the reliability, resiliency, scalability, and operational excellence of Marriott's Data and Core technology domains. This role plays a critical part in Marriott's reliability transformation, shifting the organization from reactive, centralized operations toward engineering-owned reliability grounded in modern Site Reliability Engineering (SRE) and DevSecOps practices. This leader serves as a catalyst for change, partnering closely with product, engineering, platform, architecture, and security teams to embed reliability throughout the software development lifecycle. Success is measured by measurable improvements in reliability outcomes, incident reduction, engineering adoption, and operational maturity rather than ownership of ticket queues or production support., + Own outcomes of the aligned Service Provider team inclusive of project delivery and operations + Make short term plans for the team to effectively utilize resources + Facilitate timely resolution of service delivery problems and minimizes the impact to clients Managing and Conducting Human Resources Activities, + Fosters employee commitment to providing excellent service, participates in daily stand-up meetings and models desired service behaviors in all interactions with customer and employees. + Incorporates customer satisfaction as a component of staff/operations meetings with an emphasis on generating innovative ways to continually improve results. + Sets goals and expectations for direct reports using the performance review process and holds staff accountable for performance goals. + Solicits employee feedback. + Utilizes an "open door policy" and reviews employee satisfaction results to identify and address employee problems or concerns + Promotes adherence to policies consistently, follows disciplinary procedures and documents items according to Standard and Local Operating + Procedures (SOPs and LSOPs) and supports the Peer Review Process. + Conducts annual performance appraisal with direct reports according to Standard Operating Procedures. + Champions change ensures brand and regional business initiatives are implemented and communicates follow-up actions to team as necessary. + Identifies talents of direct reports and their teams and assists with their growth and development plans. Success Measures + Increased percentage of services operating with defined SLOs and service ownership + Reduction in repeat incidents and severity of reliability events + Improved change success rates and deployment safety metrics + Increased automation and reduction of manual operational toil + Demonstrable year-over-year improvements in reliability maturity scores ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [DevSecOps: Injecting Security into Mobile CI/CD Pipelines](https://www.wearedevelopers.com/videos/273-devsecops-injecting-security-into-mobile-ci-cd-pipelines) - [The Open-source Java SDK for Multi-Cloud Development - Sandeep Pal](https://www.wearedevelopers.com/videos/2113-the-open-source-java-sdk-for-multi-cloud-development-sandeep-pal) - [We adopted DevOps and are Cloud-native, Now What?](https://www.wearedevelopers.com/videos/485-we-adopted-devops-and-are-cloud-native-now-what) - [DevSecOps culture](https://www.wearedevelopers.com/videos/783-devsecops-culture) ## Related Articles - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Top Must-Visit Developer Conferences in the US in 2026](https://www.wearedevelopers.com/magazine/679-top-must-visit-developer-conferences-in-the-us-in-2026) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Events like RSAC Get You CISOs. Developers Decide What Actually Gets Deployed.](https://www.wearedevelopers.com/magazine/693-events-like-rsac-get-you-cisos-developers-decide-what-actually-gets-deployed) - [Now is the time for industrialized software development](https://www.wearedevelopers.com/magazine/601-now-is-the-time-for-industrialized-software-development)