> Markdown version of [/jobs/ext/2653886-technical-operational-program-manager-ai-infrastructure](https://www.wearedevelopers.com/jobs/ext/2653886-technical-operational-program-manager-ai-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Technical Operational Program Manager - AI Infrastructure - **Company:** Majestic Labs LLC - **Location:** Los Altos, CA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Adobe InDesign, Artificial Intelligence, Amazon Web Services, Systems Engineering, Asana, JIRA, Microsoft Azure, Cloud Computing, Computer Clusters, Data Centers, DevOps, Distributed Computing Environment, Distributed Systems, Network Interface Controllers, Hardware Design, Network Architecture, Reliability Engineering, Systems Architecture, AI Infrastructure, Google Cloud, Build Process, Machine Learning Operations - **Published:** August 10, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=b9fe4e3203a21d84 ## About the Role * 10+ years in program/project management, including 5+ years leading technically complex, multi-stakeholder programs * Track record scaling manufacturing from prototype to volume production * Strong operational background: defining processes, metrics, and driving efficiency * Technical fluency in data center, networking, or compute infrastructure; able to engage deeply with engineers * Conceptual familiarity with AI compute, distributed training, or inference systems * Excellent written/verbal communication across technical and non-technical audiences * Proficiency with Jira, Asana, Monday.com, or similar; comfort with data/metrics dashboards, * Experience managing or scaling data center operations, cloud infrastructure (AWS, GCP, Azure), or hyperscaler / on-premises server deployments * Familiarity with HW procurement cycles, vendor management, and supply chain logistics * Background in DevOps, infrastructure engineering, or site reliability engineering (SRE) * Experience with AI/ML infrastructure, GPU clusters, or distributed computing systems * PMP, CAPM, or equivalent program management certification * Previous role in a high-growth tech or AI company Skills & Attributes * Analytical, data-driven decision-making with a metric-driven mindset (dashboards, KPIs) * Comfortable with ambiguity; able to bring clarity and proactively manage risk and contingency * Strong stakeholder influence, collaborative leadership, and attention to both detail and big-picture outcomes * Strong written and verbal communication skills ## Description We're seeking an experienced Technical Program Manager with operational expertise to lead large-scale, cross-functional engineering programs, unblock development bottlenecks, mitigate risks, and translate complex system architecture into actionable timelines for internal teams and ODM partners. This is a hybrid role that bridges technical depth with operational rigor - you'll manage complex, multi-disciplinary projects involving hardware design, procurement, deployment, network architecture, and preparation to scale. This role encompasses coordinating multiple suppliers and manufacturing partners, optimizing production workflows, and ensuring we reliably deliver our products to meet customer commitments. You'll work closely with silicon, software and systems engineers, ML teams, as well as business and operations to ensure our products are robust, reliable, and production-ready., * Program Planning & Execution: Develop roadmaps and manage end-to-end AI server deployment programs, tracking milestones, dependencies, and critical path across teams. Identify and remove development bottlenecks to ensure on-time execution. * Technical Program Management: Understand hardware specifications (silicon, interconnect, server, rack) and system architecture; participate in design reviews and translate technical decisions into program impact.. * Cross-Functional Leadership: Align engineering, product, finance, and the ODM on priorities and timelines, serving as connective tissue between technical and business stakeholders. * Supply Chain Management: Partner with Business Development to ensure timely availability of components (silicon, memory, cables and connectors, cooling, NICs, PCBs, and more). * Manufacturing Partner Coordination: Manage ODM/contract manufacturer relationships across PCB fabrication, integration, test, assembly, and QA, ensuring clarity of specs, test plans, and delivery commitments. * Operational Excellence: Build processes, metrics, and dashboards to improve product delivery and eliminate bottlenecks. * Risk & Dependency Management: Proactively identify technical, procurement, and operational risks; partner with Business Development and Systems Engineering on mitigation. * Resource Management and Timeline: Track program budgets, lead times, and hardware costs with finance and leadership to optimize resource use. * Documentation & Communication: Document key decisions and architecture; communicate status, blockers, and outcomes to all stakeholder levels. ## Related Videos - [A Founder's Journey : From Startup Chaos to Purposeful Growth](https://www.wearedevelopers.com/videos/1926-a-founder-s-journey-from-startup-chaos-to-purposeful-growth) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Improving quality with Agentic AI with Rovo Dev and Xray](https://www.wearedevelopers.com/videos/2005-improving-quality-with-agentic-ai-with-rovo-dev-and-xray) - [AI and Agility: The Dynamic Duo for Disruption](https://www.wearedevelopers.com/videos/2081-ai-and-agility-the-dynamic-duo-for-disruption) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) ## Related Articles - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again)