Senior Technical Program Manager, Cloud...

NVIDIA Ltd.
Santa Clara, CA, United States
19 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$168,000.0 - $258,750.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Confluence JIRA Big Data Cloud Computing Continuous Integration Cloud Services Software Systems Systems Integration AI Infrastructure Deep Learning Kubernetes
+2 more
Tools for Reporting Serverless Computing

Job description

NVIDIA’s deep learning platforms have made a major impact in various fields and are broadly used across leading academic institutions, start-ups, and industry, including the world’s largest Internet companies. We are seeking an experienced and talented technical program manager for NVIDIA’s DGX Cloud. We need passionate, hard-working, and creative people to help us deliver value to DGX Cloud customers.

What you’ll be doing:

As a DGX Cloud Technical Program Manager, you’ll be a key partner to our Engineering, Infrastructure, and Software teams, driving critical cloud infrastructure programs across DGX Cloud. You’ll play a pivotal role in maturing how we bring up AI capacity - strengthening process resiliency, driving automation into our PLC and acceptance workflows, and enabling early access to upcoming NVIDIA platforms. This is a dynamic, fast-paced environment where TPMs are motivated to apply fungible abilities to a range of high-impact programs.

  • Lead the end-to-end execution of NPI programs across engineering, operations, and cloud service provider (CSP) partners

  • Lead the DGX Cloud NPI Early Access Program - Enabling processes that leading to engineering teams across DGXC getting early access to critical Nvidia systems (i.e. VR / VR Ultra) in order to develop critical software and automation.

  • Drive PLC process for capacity bring-up into system-based, automated solutions - taking a baseline PLC type process and driving iterative improvements and system approaches to codifying in tooling including jira.

  • Coordinate site readiness and infrastructure bring-up activities, including networking, inventory, corp IT, and security integration

  • Partner with SW stack teams to track development, testing, and integration across product phases

  • Define and implement acceptance testing, validation workflows, and readiness gates for new platforms

  • Work closely with stakeholders to develop scalable NPI processes, tools, and dashboards

  • Drive automation efforts for break/fix workflows, telemetry enablement, and system health validation

  • Facilitate regular communication with leadership, engineering, CSP teams, and Colo partners and cultivate a culture of continuous improvement and process innovation

Requirements

  • 12+ years of technical program management experience, with a focus on infrastructure, hardware/software integration, or cloud platforms

  • Success in leading NPI or large cross-functional programs in fast-paced environments

  • Experience working with cloud service providers, large-scale data center deployments, or enterprise-scale infrastructure programs

  • Strong understanding of GPU compute, Kubernetes, CI/CD pipelines, and cloud-native services

  • Demonstrated experience building or improving product development processes and team workflows

  • Skilled in tools such as JIRA, Confluence, dashboards, and reporting tools

  • Ability to influence cross-functional teams, including HW, SW, QA, Site Ops, and Product

  • Outstanding communication and leadership skills, capable of collaborating effectively with senior collaborators

  • BS/MS in CS, EE, related technical field, or equivalent experience

Ways to stand out from the crowd:

  • Experience in launching cloud infrastructure products or large-scale hardware-software systems

  • Previous involvement in New Product Introduction (NPI), including platform bring-up and validation

  • Familiarity with AI infrastructure, or GPU-based cloud platforms

  • Experience with process automation, observability (telemetry/metrics), and health check frameworks

Benefits & conditions

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 258,750 USD for Level 4, and 200,000 USD - 322,000 USD for Level 5.

You will also be eligible for equity and benefits (https://www.nvidia.com/en-us/benefits/) .

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:33 min

Expanding enterprise search integrations with Jira, Confluence, and GitHub

Prashanth Chandrasekar Prashanth Chandrasekar · WWC 2024

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · WWC Europe 2026

5:47 min

Integrating user stories and test automation via Jira tools

Christoph Ruggenthaler · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

Videos

See all

Related articles

See all