Director - Infrastructure

REQUEST TECHNOLOGY
Houston, United States of America
7 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Senior
Compensation
$ 300K

Job location

Remote
Houston, United States of America

Tech stack

Microsoft Windows
Artificial Intelligence
Computing Platforms
Azure
Cloud Engineering
Continuous Integration
Search Technologies
Systems Integration
AI Infrastructure
Graphics Processing Unit (GPU)
System Availability
AI Platforms
Kubernetes
Azure
Machine Learning Operations
Hardware Infrastructure
Docker

Job description

  • Lead all AI environments including onpremise Graphics Processing Unit (GPU) clusters, Microsoft Azure AI and Machine Learning (ML) services, and shared AI platform components with accountability for reliability, scalability, and lifecycle management.
  • Own Azure environments hosting AI and automation workloads, including shared services such as Azure OpenAI, Azure AI Foundry, Azure AI Search, and Azure Kubernetes Service (AKS).
  • Collaborate with Cloud Engineering on landing zones, networking, subscription governance, and service onboarding, and with the AI Engineering Lead on shared platforms and operating standards.
  • Create secure, governed environments that enable rapid experimentation and development for Innovation, AI Engineering, and CGO teams.
  • Lead AI Infrastructure, AI Platform Engineering, Azure AI Engineering Operations, and Microsoft 365 (M365) Automation functions; mentor leaders and engineers and build sustainable career paths.
  • Oversee the design and deployment of shared and custom AI platforms that accelerate solution delivery while meeting security and governance standards.
  • Operationalize governance, privacy, and Responsible AI standards in partnership with Risk, Security, and Responsible AI teams.
  • Ensure platform reliability, servicelevel objectives, incident response readiness, and continuous improvement across production AI environments.
  • Manage cloud operating budgets, vendor relationships, and capacity planning across Azure services, GPU infrastructure, and AI tooling.

Requirements

  • Bachelor s degree
  • 12+ years in infrastructure, platform, or cloud engineering within complex enterprises
  • 5+ years in senior leadership roles managing managers and multiteam organizations.
  • 5+ years designing, operating, and scaling production AI and ML platforms, including MLOps, Continuous Integration and Continuous Delivery (CI/CD), InfrastructureasCode, and containerized platforms such as Kubernetes and Docker.
  • Deep expertise with enterprisescale Microsoft Azure, including AI and ML services, networking, identity, security, governance, and cost management.
  • Experience integrating and operating onpremise GPU and highperformance computing environments with cloud platforms.
  • Proven ownership of platform reliability, incident command, and servicelevel objectives in regulated, compliancesensitive environments.
  • Working knowledge of Responsible AI, AI risk management, regulatory frameworks, and compliance standards such as SOC 2 and ISO 27001.
  • Experience managing cloud spend using Financial Operations (FinOps) practices and overseeing enterprise vendor and contract relationships.
  • Strong executive presence with the ability to advise senior leaders and influence crossfunctional stakeholders.
  • Experience in legal, professional services, or similarly regulated environments preferred, including familiarity with legal technology ecosystems.

Apply for this position