> Markdown version of [/jobs/ext/2473245-token-as-a-service-technical-program-manager](https://www.wearedevelopers.com/jobs/ext/2473245-token-as-a-service-technical-program-manager). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Token-as-a-Service Technical Program Manager - **Company:** OpenAI Inc. - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $226,000.0 - $285,000.0 - **Contract:** Permanent contract - **Skills:** Systems Engineering, Computer Clusters, Data Centers, Distributed Systems, AI Infrastructure, Enterprise Integration, Machine Learning Operations - **Published:** August 21, 2026 - **Apply:** https://diversityjobs.com/main/sendform/8/8/28176/1/18001756?backUrl=%2Fcareer%2F18001756%2FToken-As-A-Service-Technical-Program-Manager-California-San-Francisco ## About the Role * 8+ years of Technical Program Management, Engineering Program Management, or Infrastructure Delivery experience. * Experience leading large-scale technical programs involving cloud, data center, networking, hardware, or distributed systems. * Strong understanding of compute infrastructure, clusters, networking, storage, and production systems. * Proven ability to drive cross-functional execution across engineering, operations, finance, and external vendors. * Experience managing executive stakeholders and communicating complex tradeoffs clearly. * Strong analytical skills with ability to reason about utilization, throughput, capacity, and operational metrics. * Comfortable operating in ambiguous, fast-scaling environments. * Strong written and verbal communication skills. * High ownership mentality with bias toward action. * Experience working with external providers, strategic partners, or hyperscalers is highly preferred. Preferred Skills * Experience with GPU clusters, AI infrastructure, or large-scale model serving environments. * Familiarity with token economics, inference capacity planning, or workload scheduling. * Experience scaling global infrastructure through third-party providers. * Background in systems engineering, networking, or hardware deployment programs. * Experience building new operational models in high-growth environments. ## Description OpenAI's Stargate and 3P Engineering teams are responsible for building and scaling the external infrastructure ecosystem that powers advanced AI systems. We work across hyperscalers, colocation providers, cloud partners, and strategic third-party operators to turn contracted capacity into production-ready compute. Our scope spans the full lifecycle of external deployments: commercial alignment, technical readiness, network integration, hardware enablement, operational readiness, and long-range scaling strategy. As OpenAI's infrastructure footprint expands globally, we need leaders who can convert complex partner environments into reliable, high-velocity capacity for training and inference workloads., We are seeking a Technical Program Manager, Token-as-a-Service (TaaS) to lead delivery of external compute capacity that directly serves OpenAI model workloads. In this role, you will own complex cross-functional programs that transform third-party infrastructure into usable tokens at scale. You will partner across engineering, capacity planning, networking, hardware, finance, product, and external providers to ensure that deployed capacity translates into real production throughput. This role sits at the intersection of infrastructure execution, systems readiness, and business impact. Success requires strong technical fluency, elite program management, and the ability to drive accountability across internal teams and external partners. This is a high-visibility role with direct impact on OpenAI's ability to scale model training and inference globally., * Lead end-to-end delivery programs that convert external infrastructure capacity into production-ready token supply. * Own readiness across compute, storage, networking, security, and operational dependencies for third-party environments. * Build integrated plans across internal engineering teams and external partners with clear milestones, owners, risks, and critical paths. * Drive launch execution for new partner regions, clusters, and capacity expansions. * Create operating mechanisms that measure deployed capacity versus usable token output. * Identify bottlenecks preventing token generation (network constraints, hardware readiness, software enablement, partner delays, etc.) and drive resolution. * Coordinate with capacity planning and finance teams to prioritize the highest ROI capacity opportunities. * Establish executive-level reporting on delivery status, risks, and token ramp forecasts. * Improve repeatability of partner onboarding, technical integration, and scaling motions. * Manage escalations across internal and external stakeholders during high-severity delivery issues. * Translate ambiguous infrastructure constraints into clear execution plans. * Help define the long-term operating model for Token-as-a-Service across Stargate and 3P ecosystems. ## Related Videos - [Building the Nervous System of AI - Michael Kagan (NVIDIA)](https://www.wearedevelopers.com/videos/2133-building-the-nervous-system-of-ai-michael-kagan-nvidia) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [The Sustainability Race: AI's Promises, Pitfalls and Potential](https://www.wearedevelopers.com/videos/100155-the-sustainability-race-ai-s-promises-pitfalls-and-potential) - [How to develop an autonomous car end-to-end: Robotic Drive and the mobility revolution](https://www.wearedevelopers.com/videos/22-how-to-develop-an-autonomous-car-end-to-end-robotic-drive-and-the-mobility-revolution) - [How to build a sovereign European AI compute infrastructure](https://www.wearedevelopers.com/videos/1102-how-to-build-a-sovereign-european-ai-compute-infrastructure) - [Building, securing and governing AI infrastructure in the Era of Agentic AI](https://www.wearedevelopers.com/videos/100129-building-securing-and-governing-ai-infrastructure-in-the-era-of-agentic-ai) ## Related Articles - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [Trustworthy AI Starts at Deployment: 5 Checks Before You Ship](https://www.wearedevelopers.com/magazine/753-trustworthy-ai-starts-at-deployment-5-checks-before-you-ship) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)