Technical Program Manager -Local AI Agents

NVIDIA Ltd.
Redmond, WA, United States
9 days ago
Apply on www.careerbuilder.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$168,000.0 - $258,750.0
Working hours
Regular working hours

Tech stack

Microsoft Windows Artificial Intelligence Application Integration Architecture C++ (Programming Language) Program Optimization Nvidia CUDA Computer Engineering Computer Literacy Desktop Computing Programming Tools Device Drivers Design of User Interfaces
+18 more
Human-Computer Interaction Python (Programming Language) Microsoft Software Open Source Technology Release Management Privacy Controls Graphics Processing Unit (GPU) Pytorch Large Language Models Multi-Agent Systems Software Security Information Technology Low Latency ONNX (Open Neural Network Exchange) Format TensorRT Virtual Agents Multiplatform Programming Languages

Job description

We seek a Senior Technical Program Manager to lead multi-functional initiatives involving local and hybrid AI agents. You will unite engineering, research, product, security, developer relations, and external collaborators behind a clear roadmap and measurable results. Your responsibilities include local model inference, agent runtimes, Windows integration, developer tools, security controls, and user experiences. Additionally, you will gather insights from open ecosystems like Hermes and OpenClaw, desktop-agent projects such as Perplexity, and domain-specific agents to guide NVIDIA’s focus on developer and user priorities.

What you’ll be doing:

  • Own the coordinated roadmap from prototype through release for local AI agent capabilities, with clear scope, achievements, owners, dependencies, and completion criteria.
  • Translate product goals and ecosystem signals into harmonized plans covering agent frameworks, MCP and tool connections, memory and skills, computer use, local inference, model routing, and application integration.
  • Partner with engineering and research teams to define evaluation and release gates for task success, latency, efficiency, memory use, power, reliability, setup time, privacy, and user trust.
  • Coordinate programs across Windows platform groups, GPU and driver groups, model and runtime groups, product security, developer relations, Microsoft, OEMs, ISVs, and open-source communities.
  • Build concise dashboards and decision forums that surface program health, technical tradeoffs, risks, and evidence; communicate clearly with engineers and senior leaders.
  • Use developer and customer feedback to improve onboarding, documentation, samples, compatibility, and adoption across supported NVIDIA systems.

Requirements

  • Bachelor’s degree in Computer Science, Computer Engineering, or a related field, or equivalent experience in practice.
  • 8+ years leading complex software, platform, systems, or AI/ML programs from concept through production release.
  • Technical proficiency in contemporary AI software stacks, encompassing model prediction, agent coordination, tool application, evaluation, and local or hybrid deployment.
  • Experience building coordinated plans across multiple engineering organizations and resolving technical dependencies without direct authority.
  • Experience defining measurable quality and release criteria, using data to make tradeoffs, and separating prototypes from validated capabilities.
  • Ability to explain architecture, risk, and program status clearly to technical teams, partners, and executives.

Ways to stand out from the crowd:

  • Hands-on experience with Python, C++, or software automation.
  • Familiarity with OpenClaw, Hermes, LangChain or similar agent frameworks; MCP; and persistent memory, skills, multi-agent, or computer-use patterns.
  • Familiarity with local inference technologies such as TensorRT-LLM, Ollama, llama.cpp, vLLM, PyTorch, ONNX Runtime, or Windows ML, plus model optimization or quantization.
  • Experience with Windows systems, CUDA or GPU acceleration, sandboxing, policy controls, privacy-aware networking, or secure credential handling.
  • Experience collaborating with Microsoft, OEMs, ISVs, open-source maintainers, or domain-solution partners., Application Integration, Artificial Intelligence (AI), Artificial Intelligence (AI) Agents, Artificial Intelligence (AI) Programming Languages, CUDA (Compute Unified Device Architecture), Communication Skills, Computer Engineering, Computer Science, Cross-Functional, Customer/Client Research, Desktop PC, Device Drivers, Documentation, Ecosystems, GPU (Graphics Processing Unit), MCP - Microsoft Certified Professional, Memory Hardware, Microsoft Product Family, Microsoft Windows Operating System, Multiplatform/Cross-Platform, OEM (Original Equipment Manufacturer), Onboarding, Open Source, Predictive Modeling, Privacy Controls, Programming Tools, Project/Program Coordination, Project/Program Management, Prototyping, Public/Media/Press/Analyst Relations, Quality Metrics, RTX, Reporting Dashboards, Risk, Technical Leadership, User Interface/Experience (UI/UX)

Benefits & conditions

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/ Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 258,750 USD for Level 4, and 200,000 USD - 322,000 USD for Level 5.

About the company

NVIDIA is the world leader in graphics processing technologies, creating innovative, industry-changing products for computing, consumer electronics, and mobile devices. NVIDIA products are transforming visually-rich applications such as video games, film production, broadcasting, industrial design, space exploration, and medical imaging. We invest in our people and our technologies, support and fund industry research around the world, and consistently deliver high-quality products. NVIDIA’s culture promotes and inspires a team of world-class employees to be at the top of their game. We’ve created an environment where talents are recognized and collaboration is valued. Our employees are shaping the world of tomorrow. . . today. We invite you to explore the opportunities available at NVIDIA to see what your future may hold.

Company Size: 10,000 employees or more

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerbuilder.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:24 min

Building client-facing AI agents for engineering teams

Alfonso Graziano Alfonso Graziano · Coffee With Developers

5:48 min

Balancing delivery latency with stream reliability and scale

Phil Cluff · LIVE

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · World Congress 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · World Congress 2026 Europe

3:37 min

Accessing API documentation and testing remote driving latency

Alexandru Ciinaru Alexandru Ciinaru +3 · World Congress 2025

Videos

See all

Related articles

See all