> Markdown version of [/jobs/ext/3537606-head-of-ai-data-center-infrastructure-platforms-and-software](https://www.wearedevelopers.com/jobs/ext/3537606-head-of-ai-data-center-infrastructure-platforms-and-software). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Head of AI Data Center Infrastructure Platforms and Software - **Company:** SUMMIT GROUP SOLUTIONS LLC - **Location:** Kirkland, WA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Computing Platforms, Software as a Service, Cloud Computing, Computer Clusters, Data Centers, Distributed Systems, Performance Tuning, Reliability Engineering, Software Engineering, AI Infrastructure, Datadog, System Availability, Containerization, Kubernetes, Infrastructure Automation Frameworks, Deployment Automation, Machine Learning Operations, Hardware Infrastructure - **Published:** October 2, 2026 - **Apply:** https://www.disabledperson.com/jobs/75658104-head-of-ai-data-center-infrastructure-platforms-and-software ## About the Role The ideal candidate combines deep technical expertise in hyperscale cloud or AI infrastructure environments with the ability to build high-performing engineering organizations in fast-scaling environments., * 15+ years of experience in infrastructure software, cloud platforms, distributed systems, or hyperscale engineering environments. * Proven experience building and scaling engineering organizations in high-growth or mission-critical environments. * Deep expertise in cloud infrastructure, Kubernetes, distributed systems, infrastructure automation, and platform engineering. * Strong understanding of AI/ML infrastructure, GPU clusters, orchestration systems, and large-scale compute environments. * Experience operating at executive or senior leadership levels within hyperscalers, cloud providers, AI infrastructure companies, telecom infrastructure organizations, or large-scale SaaS/platform companies. * Demonstrated ability to build greenfield platforms and engineering organizations in rapidly evolving environments. * Strong balance of technical depth, strategic thinking, and operational leadership. Preferred Areas of Expertise * AI infrastructure platforms and GPU cloud environments * Kubernetes and container orchestration * Infrastructure-as-code and automation frameworks * Distributed systems and large-scale cloud operations * SRE, observability, and platform reliability * Networking and high-performance infrastructure environments * Multi-region or global infrastructure deployments * Enterprise and sovereign cloud environments Leadership Profile The successful candidate will be: * Entrepreneurial and highly execution-oriented * Comfortable operating in ambiguity and building from zero to scale * Technically credible with senior engineering and infrastructure teams * Collaborative, low-ego, and highly strategic * Energized by solving complex infrastructure and scaling challenges * Motivated by the opportunity to help build a category-defining AI infrastructure company ## Description The Head of AI Infrastructure Software & Platforms will define the software and platform strategy powering a next-generation AI infrastructure ecosystem. This executive will be responsible for building the foundational software stack required to operate large-scale GPU environments supporting AI training, fine-tuning, inference, and enterprise AI workloads globally. This leader will oversee architecture, development, and operational execution across cloud infrastructure software, orchestration systems, platform services, automation frameworks, observability tooling, developer platforms, and AI infrastructure operations., Platform & Infrastructure Leadership * Define and execute the software platform strategy for a global AI cloud and GPU infrastructure environment. * Architect scalable, secure, and highly automated infrastructure platforms supporting large-scale AI workloads. * Lead development of orchestration, provisioning, scheduling, observability, and infrastructure automation systems. * Drive platform reliability, resiliency, performance optimization, and operational excellence across distributed environments. * Partner closely with hardware, networking, datacenter, and operations leaders to deliver tightly integrated AI infrastructure solutions. Engineering Organization Buildout * Build and scale a world-class software engineering and platform organization from the ground up. * Recruit, mentor, and lead engineering leaders and highly technical teams across infrastructure software, distributed systems, SRE, platform engineering, and cloud operations. * Establish engineering processes, development standards, architectural governance, and operational best practices. * Create a high-performance culture emphasizing ownership, execution velocity, innovation, and collaboration. Cloud & AI Infrastructure Strategy * Lead strategy and implementation across GPU orchestration, cluster management, Kubernetes ecosystems, virtualization/containerization, and distributed AI infrastructure. * Oversee development of platform capabilities supporting enterprise AI customers, sovereign AI deployments, and hyperscale operations. * Drive infrastructure observability, telemetry, monitoring, security, and automation initiatives. * Evaluate and integrate emerging technologies across AI infrastructure software, cloud platforms, and automation tooling. Executive Collaboration * Serve as a strategic advisor and technical partner to the CEO and executive leadership team. * Collaborate cross-functionally with infrastructure, product, operations, and business leaders to align platform strategy with commercial growth objectives. * Help shape the long-term vision, technical roadmap, and organizational scaling strategy of the company. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [The Sustainability Race: AI's Promises, Pitfalls and Potential](https://www.wearedevelopers.com/videos/100155-the-sustainability-race-ai-s-promises-pitfalls-and-potential) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [The Memory Leak That Ate Our Cluster: A Postmortem](https://www.wearedevelopers.com/videos/2057-the-memory-leak-that-ate-our-cluster-a-postmortem) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) - [Building the Nervous System of AI - Michael Kagan (NVIDIA)](https://www.wearedevelopers.com/videos/2133-building-the-nervous-system-of-ai-michael-kagan-nvidia) ## Related Articles - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)