Senior Technical Program Manager - Capacity

NVIDIA Ltd.
Santa Clara, CA, United States
about 2 months ago
Apply on www.juju.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$200,000.0 - $322,000.0
Working hours
Regular working hours
Job source

Tech stack

Data Centers Cloud Services Information Technology Data Analytics Hardware Infrastructure

Job description

Hardware Infrastructure is seeking a Technical Program Manager to lead Infrastructure Capacity Management programs and workstreams. Given this Infrastructure directly supports our near-term and long-term chip roadmap, it must be highly reliable, performant and efficient for our internal users. This is a fast paced and evolving landscape that requires a TPM to guide engineering roadmaps to be delivered with high quality outcomes and a strong foundation of operational excellence. They will partner both internally within Hardware Infrastructure and externally with senior management and our HW Engineering partners to manage capacity operations and scale processes supporting our next stage of growth. They will develop and standardize planning, reporting and execution methodologies and metrics to enable meeting the challenging objectives.

What You’ll Be Doing:

  • Own end-to-end capacity management strategy and execution for EDA Farm, including server procurement, vendor negotiations, server capacity allocation, data center space & power, delivering measurable efficiency and cost optimization across the organization

  • Identify and help drive implementation of improvements to EDA Farm infrastructure tooling, automation, and workflows that accelerate server provisioning, reduce manual overhead, and scale capacity management operations

  • Drive capacity and procurement initiatives using agile program methodology, aligning planning, prioritization, and delivery across engineering, procurement, and vendor partner teams

  • Build and maintain a data-driven capacity model, using metrics and business objectives to improve farm utilization, procurement performance, and vendor SLA consistency - turning insights into actionable cost and capacity optimizations

  • Create clear, consistent communication channels that give customers at every level real-time insight into farm capacity health, procurement timelines, supply chain risks, and mitigation plans

  • Act as a primary technical and strategic partner between engineering, procurement, finance, and hardware vendors to ensure EDA Farm capacity optimally meets the demands of design and verification teams

Requirements

  • B.S. (or equivalent experience) in Electrical Engineering, Computer Science or a related technical field

  • 12+ years of proven experience across Capacity Engineering, Capacity Management and/or Technical Program Management roles within the Capacity space

  • Experience working with large scale infrastructure with various CPU/GPU architectures both on-prem and cloud

  • Exceptional communication and presentation skills for diverse technical and non-technical audiences

  • Proactive in identifying and implementing positive changes in both system and process design in a fast-paced environment

Ways To Stand Out From The Crowd:

  • Deep experience leading Capacity Operations & Management work streams

  • Prior experience procuring Data Center Hardware and Public Cloud Services

Benefits & conditions

NVIDIA offers highly competitive salaries and a comprehensive benefits package. We have some of the most forward-thinking and hardworking people in the world on our team and our collaborative talent continues to drive NVIDIA’s growth. We are seeking creative and independent engineers with real passion for technology!

LI-Hybrid

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 200,000 USD - 322,000 USD.

You will also be eligible for equity and benefits (https://www.nvidia.com/en-us/benefits/) .

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.juju.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:12 min

Addressing the competitive landscape of specialized hardware demands

Hazal Mestci +1 · Coffee With Developers

51 sec

Repurposing hardware and operating underwater data centers

Chris Heilmann +1 · LIVE

2:41 min

Transitioning artificial intelligence infrastructure into scalable commodity cloud services

juarezjunior juarezjunior · World Congress 2024

1:32 min

Structuring platforms for new services and data analytics

Nevelina Aleksandrova · LIVE

1:29 min

Tech infrastructure capacity and AI product innovations

4:03 min

Managing massive power consumption scaling in AI data centers

Stephan Gillich Stephan Gillich +3 · World Congress 2024

Videos

See all

Related articles

See all