IT leader

Mitchell International, Inc.
United States
8 days ago
Apply on jobs.military.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$130,800.0 - $193,000.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Microsoft Azure Cloud Computing Configuration Management Databases Data Centers DevOps Disaster Recovery Domain Name System (DNS) Identity and Access Management Information Technology Operations Reliability Engineering Software Engineering
+2 more
Software Vulnerability Management Information Technology

Requirements

  • Bachelor’s Degree in Computer Science, Software Engineering, or related field\n
  • 10+ years of experience in IT Operations, Cloud Operations, Infrastructure Operations, or Site Reliability Engineering, including 5+ years in leadership roles.\n
  • Experience leading enterprise-scale AWS and/or Azure environments with accountability for reliability, availability, security, operational performance, and service delivery.\n
  • Strong background in ITSM, operational excellence, and operational governance, including incident, problem, change, configuration, and service management.\n
  • Experience establishing service ownership, operational readiness, and service transition processes for production services and cloud platforms.\n
  • Experience managing strategic vendors, managed service providers, contracts, service reviews, and SLA accountability.\n
  • Demonstrated success improving SLAs, SLOs, KPIs, operational maturity, and service performance.\n
  • Experience leading FinOps, automation, infrastructure-as-code adoption, and continuous improvement initiatives.\n
  • Strong understanding of resiliency, disaster recovery, business continuity, security, compliance, risk management, and executive stakeholder leadership.\n, * AWS and/or Azure certifications (Solutions Architect, SysOps Administrator, DevOps Engineer, or equivalent).\n
  • ITIL certification or formal service management training.\n
  • Experience operating in regulated or highly compliant environments.\n
  • Experience leading large-scale cloud migration, modernization, or data center exit initiatives.\n

\n \n

  • Experience with cloud governance, observability, CSPM, CMDB, vulnerability management, IAM integration, DNS operations, messaging dependencies, and policy compliance\n

Benefits & conditions

The Director of Cloud Operations is a senior IT leader responsible for the reliability, security, cost efficiency, operational readiness, and continuous improvement of the enterprise cloud environment across AWS and Azure. This leader owns the day-to-day operation of cloud infrastructure and services, ensuring platforms are resilient, secure, supportable, and operationally ready throughout their lifecycle.\n \n The Director leads the Cloud Operations organization and is accountable for service reliability, operational governance, vendor performance, cloud financial management (FinOps), and the successful transition of new cloud services into production support. The role establishes the standards, processes, and operating model required to deliver scalable, resilient, and cost-effective cloud services while partnering closely with engineering, architecture, security, and business teams.\n \n \nKey Responsibilities:\n \n \nLeadership & Team Management\n \n \n

  • Lead, mentor, and develop a high-performing cloud operations organization.\n
  • Foster a culture of accountability, ownership, operational excellence, and continuous improvement.\n
  • Establish performance expectations, career development plans, and operational readiness standards.\n
  • Build organizational capabilities and succession depth across AWS and Azure operations.\n

\n \nCloud Operations & Reliability (AWS & Azure)\n \n \n

  • Own the day-to-day operation of AWS and Azure environments, including availability, performance, patching, backup, and recovery.\n
  • Ensure disciplined execution of operational processes including monitoring, incident response, configuration management, and patch management.\n
  • Lead incident, problem, and change management processes with a focus on root cause analysis and prevention.\n
  • Establish observability, alerting, automated remediation, and operational health monitoring across the cloud environment.\n
  • Lead disaster recovery, business continuity, and resiliency planning for cloud services.\n
  • Establish capacity, performance, configuration, and service management practices, including service ownership, escalation models, asset management, and CMDB accuracy.\n

\n \n

  • Establish and maintain operational support coverage, on-call processes, escalation procedures, and incident response readiness.\n

\n \nOperational Readiness & Service Transition\n \n \n

  • Establish operational readiness and acceptance requirements for cloud workloads entering production, including monitoring, backup, recovery, documentation, runbooks, support ownership, identity and access validation, DNS readiness, messaging dependencies, vulnerability status, and security control evidence.\n

\n \n

  • Lead service transition activities for new cloud services, migrations, and major platform changes.\n

\n \n

  • Ensure support models, escalation paths, operational procedures, security control validation, IAM integration, DNS readiness, and messaging dependencies are in place prior to production go-live.\n
  • Partner with application, engineering, architecture, security, IAM, messaging, DNS, network, and compliance teams to identify and mitigate operational, control, dependency, and supportability risks before deployment.\n

\n \nGovernance, Security & Compliance\n \n \n

  • Partner with Architecture, Engineering, Security teams to operationalize approved cloud governance, security controls, vulnerability remediation, access requirements, service dependencies, and compliance obligations.\n
  • Ensure adherence to regulatory, audit, and internal control requirements across cloud environments.\n
  • Operate cloud configuration standards, tagging strategies, access governance, and policy compliance in alignment with enterprise IAM, security architecture, and governance requirements.\n

\n \n

  • Identify and manage operational risks, including technical debt, lifecycle risks, resiliency gaps, supportability concerns, IAM dependencies, DNS dependencies, messaging dependencies, and security remediation exposure.\n

\n \nVendor & Managed Services Management\n \n \n

  • Own strategic relationships with cloud providers and managed service partners.\n
  • Manage vendor performance, contracts, service reviews, escalations, and accountability to SLAs and service commitments.\n
  • Evaluate, negotiate, and optimize managed service and tooling engagements for value and outcomes.\n
  • Ensure clear ownership boundaries, support responsibilities, and transition processes between internal teams and service providers.\n
  • Drive continual service improvement initiatives with strategic partners to improve reliability, efficiency, and customer experience.\n

\n \nSLA, Metrics & FinOps (Cloud Financial Management)\n \n \n

  • Define and report SLAs, SLOs, KPIs, service health metrics, and operational scorecards.\n

\n \n

  • Establish dashboards and service reviews that provide visibility into reliability, customer experience, cost, operational risks, security remediation, and improvement opportunities.\n

\n \n

  • Lead the cloud FinOps practice, including budgeting, forecasting, cost allocation, optimization, and financial accountability.\n
  • Drive continuous cost optimization through rightsizing, commitment management, and waste elimination across AWS and Azure.\n
  • Partner with Finance on cloud spend planning, reporting, and variance management.\n
  • Use data-driven insights to prioritize investments, operational improvements, and technology decisions.\n

\n \n

  • Measure and improve customer experience through service quality metrics, stakeholder feedback, and operational service reviews.\n

\n \nAutomation & Modernization\n \n \n

  • Define and drive the cloud automation strategy, including automated provisioning, remediation, self-healing capabilities, operational workflows, and lifecycle management processes.\n
  • Partner with Engineering, Security, and Architecture teams to operationalize infrastructure-as-code standards and improve operational supportability.\n
  • Measure automation adoption and operational efficiency gains, with a focus on reducing manual effort and improving reliability.\n

\n \nStrategic & Cross-Functional Responsibilities\n \n \n

  • Provide operational input into cloud architecture and engineering decisions, including reliability, resiliency, supportability, and lifecycle management requirements.\n
  • Collaborate with security, network, identity, application, finance, and business teams to deliver integrated outcomes.\n
  • Maintain operational documentation, standards, runbooks, and support playbooks.\n
  • Partner with business and technology leaders to ensure cloud services align with operational requirements, support models, and enterprise standards.\n
  • Lead regular operational and service reviews to assess performance, risks, trends, and improvement opportunities.\n, We’re committed to supporting your ultimate well-being through our total compensation package offerings that support your health, wealth and self. These offerings include Medical, Dental, Vision, Health Savings Accounts / Flexible Spending Accounts, Life and AD&D Insurance, 401(k), Tuition Reimbursement, and an array of resources that encourage a lifetime of healthier living. Benefits eligibility may differ depending on full-time or part-time status. Compensation depends on the applicable US geographic market. The expected base pay for this position ranges from $130,800 - $193,000 annually, and will be based on a number of additional factors including skills, experience, and education. \n \n The Company is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, religion, color, national origin, gender, gender identity, sexual orientation, age, status as a protected veteran, among other things, or status as a qualified individual with disability. \n \n Don’t meet every single requirement? Studies have shown that women and underrepresented minorities are less likely to apply to jobs unless they meet every single qualification. We are dedicated to building a diverse, inclusive, and authentic workplace, so if you’re excited about this role but your past experience doesn’t align perfectly with every qualification in the job description, we encourage you to apply anyway. You may be just the right candidate for this or other roles.\n \n #LI-FP1\n \n #LI-Remote\n \n $130800 - $193000 annually\n \n PI286707335”, “hiringOrganization”: {“@type”: “Organization”, “name”: “Mitchell International”}, “jobLocation”: {“address”: {“addressCountry”: “United States”, “streetAddress”: “Not specified”, “@type”: “PostalAddress”, “postalCode”: “92121”, “addressLocality”: “San Diego”, “addressRegion”: “California - CA”}, “@type”: “Place”}, “industry”: “”, “identifier”: {“@type”: “PropertyValue”, “name”: “Mitchell International”, “value”: “286707335”}, “baseSalary”: {“@type”: “MonetaryAmount”, “currency”: “USD”, “value”: {“@type”: “QuantitativeValue”, “value”: “Competitive”, “unitText”

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jobs.military.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:46 min

Navigating a career in cloud transformation consulting

Piet Van Dongen · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

51 sec

Repurposing hardware and operating underwater data centers

Chris Heilmann +1 · LIVE

1:06 min

Outline of free tools for Microsoft Azure

Radu Vunvulea Radu Vunvulea · World Congress 2022

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all