Staff Technical Program Manager, Deployments

Crusoe's Inc
San Francisco, CA, United States
20 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Compensation
$200,000.0 - $240,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Data Analysis BIOS Cloud Computing Firmware AI Infrastructure Graphics Processing Unit (GPU) Data Analytics

Job description

  • Own the infrastructure deployment for new sites or site expansions end-to-end: chip vendor and OEM dependencies, architecture updates, cloud foundations work, commissioning gate framework definition, and first customer cluster delivery as the success metric.
  • Lead Deployment Phase 0 on the Cloud TPM side: define firmware version targets and DOCA targets before kickoff, set commissioning gate criteria, and define the DRI matrix.
  • Manage compounding cross-SKU dependencies where active production programs (e.g. B200/GB300/VR) are running in parallel with capacity expansion projects, and prevent them from competing for the same Engineering pool without a plan.

Technical Leadership

  • Owns real-time execution dashboards; delivers crisp, data-driven executive updates that surface decision elements without requiring follow-up
  • Governs cross-organizational dependencies without waiting for escalation authority

Organizational Influence and TPM Function Development

  • Coach more junior TPMs on technical depth, risk identification, and executive communication.
  • Actively drives AI tool integration across their programs; identifies where AI materially improves program tracking, risk detection, and executive communication

What You’ll Bring to the Team:

Technical Foundation

  • Deep, working fluency with GPU architecture across SKU generations, firmware lifecycle (DOCA, driver stacks, BIOS/BMC), compute orchestration, SDN, storage, networking (leaf-spine topology, ZTP, fabric commissioning), and monitoring/observability
  • Direct hardware partner engagement: personal ownership of NVIDIA or OEM certification and validation timelines, not coordination feeding into someone else’s relationship.
  • Active daily use of AI tools to drive program-level outcomes: risk detection, dependency mapping, data analysis, and executive communication, not just personal productivity.

Requirements

  • 10+ years as a Technical Program Manager with a track record of owning infrastructure deployment programs end-to-end at a hyperscaler, GPU cloud provider, or AI infrastructure company, ideally with direct experience in large-scale capacity expansion projects.
  • Proven ability to define deployment engagement models from scratch, not just operate within existing frameworks, and make them stick across engineering organizations that didn’t ask for them.
  • Track record of driving cross-organizational alignment at VP/SVP level without formal authority, including building durable alignment on programs that fall in the cracks between teams.
  • Exceptional written and verbal communication for delivering clear, data-driven, decision-oriented updates to executive stakeholders.

Bonus Points:

  • Experience defining or substantially redesigning a commissioning gate framework for site deployments.
  • Experience coaching or developing more junior TPMs in technical depth and program execution.

Benefits & conditions

Pulled from the full job description

  • Tuition reimbursement
  • Paid parental leave
  • Parental leave
  • 401(k) matching
  • Paid time off
  • Vision insurance
  • Health savings account, * Competitive compensation and equity packages
  • Restricted Stock Units
  • Paid time off, paid holidays & leave of absence programs
  • Comprehensive health, dental & vision insurance
  • Employer contributions to HSA account
  • Paid parental leave
  • Paid life insurance, short-term and long-term disability
  • Professional development & tuition reimbursement
  • Mental health & wellness support
  • Commuter benefits (parking & transit)
  • Cell phone stipend
  • 401(k) Retirement plan with company match up to 4% of salary
  • Volunteer time off
  • Global travel insurance & emergency assistance
  • Daily meals allowance
  • Additional perks & programs specific to location

Compensation Range

Compensation will be paid in the range of $200,000 to $240,000.. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant’s knowledge, education, and abilities, as well as internal equity and alignment with market data.

About the company

Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack - from electrons to tokens - to power the world’s most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.

We’re in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We’re solving that - with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.

We’re looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.

If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.

About This Role:

Crusoe is the world’s first vertically integrated, sustainable AI cloud. We build and operate GPU infrastructure powered by clean energy, from data center design through IaaS products to managed inference at scale, enabling AI-native companies to run demanding workloads without compromising on sustainability or reliability. Crusoe Cloud is 1,400 people and growing. The TPM frameworks are still being built, which means there is a real opportunity to shape how the function operates rather than inherit how it already works.

We are hiring a Staff TPM who will own deployment programs. You will be helping us build out new sites, capacity expansion, or help deploy modular (Spark) data centers. Our TPMs define and drive the entire deployment engagement model from chip vendor engagement through first customer cluster delivery.

If you have spent your career running infrastructure deployment programs at a hyperscaler or a leading neocloud, understand GPU architecture, and have shaped deployment frameworks rather than just operated within them, this is the role for you.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:48 min

Automating exploratory data analysis within training pipelines

Dora Petrella · WWC 2023

41 sec

Massive client data loss and bio-digital storage

Chris Heilmann +1 · LIVE

2:19 min

Orchestrating over-the-air firmware updates for vehicle modules

Denis Grahovac · WWC 2021

1:06 min

Developer experience and project variety at scale

Alexandra Petri · WWC 2023

1:36 min

Performing exploratory data analysis to uncover underlying patterns

Julian Joseph · LIVE

Videos

See all

Related articles

See all