> Markdown version of [/jobs/ext/516750-project-manager-cluster-deployment](https://www.wearedevelopers.com/jobs/ext/516750-project-manager-cluster-deployment). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Project Manager, Cluster Deployment - **Company:** STN, inc. - **Location:** Pleasanton, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $125,000.0 - $165,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Asana, Computer Clusters, Data Centers, RAID, InfiniBand, Microsoft Project, PRINCE2, AI Infrastructure, Smartsheet, Information Technology - **Published:** June 13, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=e11c13bb4122f90a ## About the Role Do you have experience in Vendor relationship management?, * 5+ years PM experience delivering data center, colocation, or large-scale infrastructure projects * Demonstrated ownership of multi-million dollar deployments with vendor, customer, and internal stakeholder coordination * Strong working knowledge of data center physical infrastructure (power, cooling, cabling, racks) and IT infrastructure (compute, storage, networking) * Proficient with project tooling (MS Project, Smartsheet, Asana, or equivalent) and standard collaboration stack (M365, Slack) * Excellent written communication; comfortable producing exec-ready status and customer-facing updates Preferred * Direct experience deploying GPU clusters (NVIDIA HGX/DGX, B200/B300/GB300) or HPC/AI infrastructure * Experience with modular data center (MDC) builds or greenfield site delivery * Familiarity with InfiniBand/RoCE fabrics, liquid cooling, and high-density power designs * PMP, PRINCE2, or equivalent certification * Background working with hyperscaler, AI lab, or enterprise AI customers Success Metrics (first 12 months) * On-time go-live across assigned deployments vs. contracted dates * Budget variance within ±5% of approved plan * Customer acceptance and CSAT at handoff * Reduction in deployment cycle time across repeat build patterns * Quality of project documentation and post-mortem outputs Travel Up to 40% to deployment sites and customer locations. ## Description Deployment Execution * Own the master project plan, RAID log, and critical path for each cluster deployment from PO through customer acceptance * Drive site readiness across power, cooling, structured cabling, network, and physical security * Manage hardware logistics: GPU/server/networking receipt, staging, rack-and-stack, burn-in, and acceptance testing * Coordinate cross-functional teams across Architecture, Engineering, Managed Services, and Field Operations Vendor and Partner Management * Manage MDC, EPC, OEM, and colocation vendor schedules and deliverables (e.g., Mod42, Collier, CoreSite, Equinix) * Run weekly vendor scorecards; escalate slips, change orders, and dependency risks * Coordinate freight, cargo insurance, customs (where applicable), and high-value GPU shipment receiving Customer delivery * Serve as primary technical PM interface to customers from kickoff through go-live and handoff to CSM * Maintain customer-facing schedules, status reports, and milestone communications * Coordinate acceptance testing, SLA baseline validation, and project closeout documentation Financial and contractual * Track project budget, capex/opex burn, vendor invoicing, and milestone-based payments against the service order or MSA * Flag scope changes and drive change-order process with Sales, Legal, and Finance * Partner with Finance on revenue recognition triggers (deposit, MRC start, true-up) Governance and reporting * Produce weekly executive status (Sabur/Tom/Armar) and customer status reports * Maintain deployment runbooks, lessons learned, and template artifacts to scale future builds ## Related Videos - [A Founder's Journey : From Startup Chaos to Purposeful Growth](https://www.wearedevelopers.com/videos/1926-a-founder-s-journey-from-startup-chaos-to-purposeful-growth) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) - [AI Factories at Scale](https://www.wearedevelopers.com/videos/1139-ai-factories-at-scale) ## Related Articles - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Why Attend a Developer Event in 2026?](https://www.wearedevelopers.com/magazine/688-why-attend-a-developer-event-in-2026) - [5 Reasons Why Attending Conferences in 2026 Matters More Than You Think](https://www.wearedevelopers.com/magazine/692-5-reasons-why-attending-conferences-in-2026-matters-more-than-you-think) - [How to Turn Community Events Into a Powerful AI GTM Engine: The Daytona Playbook](https://www.wearedevelopers.com/magazine/732-how-to-turn-community-events-into-a-powerful-ai-gtm-engine-the-daytona-playbook) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)