Technical Program Manager -Local AI Agents

NVIDIA Corporation
Santa Clara, CA, United States
9 days ago
Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$168,000.0 - $258,750.0
Working hours
Regular working hours

Tech stack

Microsoft Windows Artificial Intelligence Application Integration Architecture C++ (Programming Language) Program Optimization Nvidia CUDA Computer Engineering Computer Literacy Programming Tools Python (Programming Language) Open Source Technology Release Management
+9 more
Pytorch Large Language Models Multi-Agent Systems Software Security Information Technology Low Latency ONNX (Open Neural Network Exchange) Format TensorRT Virtual Agents

Job description

We seek a Senior Technical Program Manager to lead multi-functional initiatives involving local and hybrid AI agents. You will unite engineering, research, product, security, developer relations, and external collaborators behind a clear roadmap and measurable results. Your responsibilities include local model inference, agent runtimes, Windows integration, developer tools, security controls, and user experiences. Additionally, you will gather insights from open ecosystems like Hermes and OpenClaw, desktop-agent projects such as Perplexity, and domain-specific agents to guide NVIDIA’s focus on developer and user priorities.

What you’ll be doing:

  • Own the coordinated roadmap from prototype through release for local AI agent capabilities, with clear scope, achievements, owners, dependencies, and completion criteria.
  • Translate product goals and ecosystem signals into harmonized plans covering agent frameworks, MCP and tool connections, memory and skills, computer use, local inference, model routing, and application integration.
  • Partner with engineering and research teams to define evaluation and release gates for task success, latency, efficiency, memory use, power, reliability, setup time, privacy, and user trust.
  • Coordinate programs across Windows platform groups, GPU and driver groups, model and runtime groups, product security, developer relations, Microsoft, OEMs, ISVs, and open-source communities.
  • Build concise dashboards and decision forums that surface program health, technical tradeoffs, risks, and evidence; communicate clearly with engineers and senior leaders.
  • Use developer and customer feedback to improve onboarding, documentation, samples, compatibility, and adoption across supported NVIDIA systems.

Requirements

  • Bachelor’s degree in Computer Science, Computer Engineering, or a related field, or equivalent experience in practice.
  • 8+ years leading complex software, platform, systems, or AI/ML programs from concept through production release.
  • Technical proficiency in contemporary AI software stacks, encompassing model prediction, agent coordination, tool application, evaluation, and local or hybrid deployment.
  • Experience building coordinated plans across multiple engineering organizations and resolving technical dependencies without direct authority.
  • Experience defining measurable quality and release criteria, using data to make tradeoffs, and separating prototypes from validated capabilities.
  • Ability to explain architecture, risk, and program status clearly to technical teams, partners, and executives.

Ways to stand out from the crowd:

  • Hands-on experience with Python, C++, or software automation.
  • Familiarity with OpenClaw, Hermes, LangChain or similar agent frameworks; MCP; and persistent memory, skills, multi-agent, or computer-use patterns.
  • Familiarity with local inference technologies such as TensorRT-LLM, Ollama, llama.cpp, vLLM, PyTorch, ONNX Runtime, or Windows ML, plus model optimization or quantization.
  • Experience with Windows systems, CUDA or GPU acceleration, sandboxing, policy controls, privacy-aware networking, or secure credential handling.
  • Experience collaborating with Microsoft, OEMs, ISVs, open-source maintainers, or domain-solution partners.

Benefits & conditions

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 258,750 USD for Level 4, and 200,000 USD - 322,000 USD for Level 5.

You will also be eligible for equity and benefits.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · World Congress 2026 Europe

5:48 min

Balancing delivery latency with stream reliability and scale

Phil Cluff · LIVE

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · World Congress 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

1:24 min

Building client-facing AI agents for engineering teams

Alfonso Graziano Alfonso Graziano · Coffee With Developers

3:37 min

Accessing API documentation and testing remote driving latency

Alexandru Ciinaru Alexandru Ciinaru +3 · World Congress 2025

Videos

See all

Related articles

See all