> Markdown version of [/jobs/ext/278152-senior-technical-program-manager](https://www.wearedevelopers.com/jobs/ext/278152-senior-technical-program-manager). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Technical Program Manager - **Company:** Crusoe's Inc - **Location:** Sunnyvale, CA, United States - **Experience:** Expert - **Salary:** $161,700.0 - $196,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Computing Platforms, Systems Engineering, Microsoft Azure, BIOS, Cloud Computing, Nvidia CUDA, Computer Engineering, Data Centers, Embedded Software, Firmware, Hardware Platform Interface, Infrastructure as a Service (IaaS), Cloud Services, AI Infrastructure, Pytorch, Information Technology, Oracle Cloud Infrastructure - **Published:** May 19, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=5952c47aab129660 ## About the Role Do you have experience in Cloud-based systems?, Do you have a Master's degree?, * AI Infrastructure Technical Depth: Genuine hands-on knowledge of the AI infrastructure stack - GPU firmware, drivers, BIOS/BMC, CUDA or ROCm stacks, OS configurations, and the tooling required to bring a cluster to production. Ability to engage credibly with firmware, hardware, networking, and compute engineers without a translator. * Commissioning Fluency: Familiarity with GPU cluster commissioning - understanding of the distinct workstreams (Compute, SDN, ZTP, Firmware, Network, Storage), commissioning gate definitions, and what it means to take ownership of a cluster from RFN through GA. * Experience at Hyperscalers or Neoclouds: 5-10 years of experience as a Technical Program Manager, with meaningful tenure at a hyperscaler (AWS, Azure, Google, Meta, OCI) or neocloud (CoreWeave, Lambda Labs) in a hardware, firmware, compute platform, or IaaS delivery role. Must have been on the build side - not customer-facing or enterprise IT delivery. * AI Tool Integration: Active daily use of AI tools to enhance program execution, tracking, and communication. * Execution Rigor: Demonstrated ability to establish program predictability in ambiguous environments. Builds lightweight structure that engineering teams trust and adopt. * Communication: Clear written and verbal communication for delivering program status, risks, and dependencies to engineering leads and product managers. Bonus Points * Bachelor's or Master's degree in Engineering, Computer Science, or a related technical field * Hands-on engineering background - hardware engineering, systems engineering, silicon validation, embedded software, or firmware * Experience with GPU NPI programs, firmware deployment across compute fleets, or datacenter commissioning * Experience at a hyper-growth company ## Description Crusoe's Cloud Product team is hiring a Senior Technical Program Manager to own and drive technical programs across our AI IaaS platform, aligning Product and Engineering on delivery. This role requires a hands-on individual contributor with genuine technical depth in AI infrastructure, comfortable operating in ambiguous, fast-moving environments and capable of building execution structure where little exists. Our vision is the easiest-to-use AI purpose-built cloud. We offer IaaS products, letting AI/ML engineers focus on AI model frameworks (PyTorch, Ray) and compute stacks (CUDA, ROCm) while Crusoe manages the underlying complexity. Our platform standardizes firmware/OS bundles and automates component orchestration for consistent, scalable infrastructure. We also offer AI Managed Services, like SLA-bounded Managed Inference. The TPM connects engineering, product, procurement, and data center operations to deliver a reliable platform where customers run AI workloads without managing low-level system details. At this level, TPM engagement focuses on defined technical programs and workstreams within larger cross-functional initiatives: GPU cluster commissioning, feature delivery within IaaS products, NPI workstream ownership, and coordination across hardware and software dependencies. The ideal candidate is engineer-rooted with hands-on technical depth in compute infrastructure, firmware, or hardware platform delivery. You will own defined programs end-to-end, build lightweight execution frameworks, govern dependencies within your scope, and grow into broader program ownership over time. What You'll Be Working On: * Workstream & Program Ownership: Own delivery of defined feature sets or commissioning workstreams within larger IaaS and NPI programs. Drive execution from kickoff through GA or handoff milestone. * Hands-On Execution: Personally manage technically complex programs within your scope. Identify blockers early, intervene to correct course, and escalate cross-organizational issues promptly when they exceed your authority. * Commissioning Leadership: Own the Cloud TPM side of GPU cluster commissioning for assigned deployments - define RFN acceptance criteria, commissioning readiness requirements, and Go/No-Go criteria. Lead Phase 4.0 commissioning activities including network fabric, storage provisioning, GPU validation, burn-in, and monitoring integration. * Risk & Dependency Governance: Proactively identify technical and organizational risks within your programs. Maintain dependency logs, surface blockers before scheduled reviews, and drive issues to resolution. * Execution Frameworks: Build lightweight, fit-for-purpose execution structures for your programs - standups, dashboards, status updates, and risk logs. Establish predictability within your workstream. * External Partner Coordination: Coordinate with external dependencies such as hardware partners, ODMs, or data center operations teams on assigned programs. Support certification, validation, and infrastructure stack alignment under guidance of senior TPMs. ## Related Videos - [10M Data Records Lost, Underwater Computing, and Psychedelic Fish - Matthias Geniar](https://www.wearedevelopers.com/videos/1908-10m-data-records-lost-underwater-computing-and-psychedelic-fish-matthias-geniar) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Playing Pong on a shoulder press machine](https://www.wearedevelopers.com/videos/100140-playing-pong-on-a-shoulder-press-machine) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [How to Submit CFPs and Get into Public Speaking - Moran Weber](https://www.wearedevelopers.com/videos/2112-how-to-submit-cfps-and-get-into-public-speaking-moran-weber) - [The Software Engineer 2030: From Coder To AI Orchestrator? - Patrick Schnell](https://www.wearedevelopers.com/videos/1825-the-software-engineer-2030-from-coder-to-ai-orchestrator-patrick-schnell) ## Related Articles - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this)