Technical Program Manager -Local AI Agents
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+18 more
Job description
We seek a Senior Technical Program Manager to lead multi-functional initiatives involving local and hybrid AI agents. You will unite engineering, research, product, security, developer relations, and external collaborators behind a clear roadmap and measurable results. Your responsibilities include local model inference, agent runtimes, Windows integration, developer tools, security controls, and user experiences. Additionally, you will gather insights from open ecosystems like Hermes and OpenClaw, desktop-agent projects such as Perplexity, and domain-specific agents to guide NVIDIA’s focus on developer and user priorities.
What you’ll be doing:
- Own the coordinated roadmap from prototype through release for local AI agent capabilities, with clear scope, achievements, owners, dependencies, and completion criteria.
- Translate product goals and ecosystem signals into harmonized plans covering agent frameworks, MCP and tool connections, memory and skills, computer use, local inference, model routing, and application integration.
- Partner with engineering and research teams to define evaluation and release gates for task success, latency, efficiency, memory use, power, reliability, setup time, privacy, and user trust.
- Coordinate programs across Windows platform groups, GPU and driver groups, model and runtime groups, product security, developer relations, Microsoft, OEMs, ISVs, and open-source communities.
- Build concise dashboards and decision forums that surface program health, technical tradeoffs, risks, and evidence; communicate clearly with engineers and senior leaders.
- Use developer and customer feedback to improve onboarding, documentation, samples, compatibility, and adoption across supported NVIDIA systems.
Requirements
- Bachelor’s degree in Computer Science, Computer Engineering, or a related field, or equivalent experience in practice.
- 8+ years leading complex software, platform, systems, or AI/ML programs from concept through production release.
- Technical proficiency in contemporary AI software stacks, encompassing model prediction, agent coordination, tool application, evaluation, and local or hybrid deployment.
- Experience building coordinated plans across multiple engineering organizations and resolving technical dependencies without direct authority.
- Experience defining measurable quality and release criteria, using data to make tradeoffs, and separating prototypes from validated capabilities.
- Ability to explain architecture, risk, and program status clearly to technical teams, partners, and executives.
Ways to stand out from the crowd:
- Hands-on experience with Python, C++, or software automation.
- Familiarity with OpenClaw, Hermes, LangChain or similar agent frameworks; MCP; and persistent memory, skills, multi-agent, or computer-use patterns.
- Familiarity with local inference technologies such as TensorRT-LLM, Ollama, llama.cpp, vLLM, PyTorch, ONNX Runtime, or Windows ML, plus model optimization or quantization.
- Experience with Windows systems, CUDA or GPU acceleration, sandboxing, policy controls, privacy-aware networking, or secure credential handling.
- Experience collaborating with Microsoft, OEMs, ISVs, open-source maintainers, or domain-solution partners., Application Integration, Artificial Intelligence (AI), Artificial Intelligence (AI) Agents, Artificial Intelligence (AI) Programming Languages, CUDA (Compute Unified Device Architecture), Communication Skills, Computer Engineering, Computer Science, Cross-Functional, Customer/Client Research, Desktop PC, Device Drivers, Documentation, Ecosystems, GPU (Graphics Processing Unit), MCP - Microsoft Certified Professional, Memory Hardware, Microsoft Product Family, Microsoft Windows Operating System, Multiplatform/Cross-Platform, OEM (Original Equipment Manufacturer), Onboarding, Open Source, Predictive Modeling, Privacy Controls, Programming Tools, Project/Program Coordination, Project/Program Management, Prototyping, Public/Media/Press/Analyst Relations, Quality Metrics, RTX, Reporting Dashboards, Risk, Technical Leadership, User Interface/Experience (UI/UX)
Benefits & conditions
Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/ Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 258,750 USD for Level 4, and 200,000 USD - 322,000 USD for Level 5.
About the company
NVIDIA is the world leader in graphics processing technologies, creating innovative, industry-changing products for computing, consumer electronics, and mobile devices. NVIDIA products are transforming visually-rich applications such as video games, film production, broadcasting, industrial design, space exploration, and medical imaging. We invest in our people and our technologies, support and fund industry research around the world, and consistently deliver high-quality products. NVIDIA’s culture promotes and inspires a team of world-class employees to be at the top of their game. We’ve created an environment where talents are recognized and collaboration is valued. Our employees are shaping the world of tomorrow. . . today. We invite you to explore the opportunities available at NVIDIA to see what your future may hold.
Company Size: 10,000 employees or more
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
What is Agentic Programming and Why Should Developers Care?
Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?
MLOps – What’s the deal behind it?
Everything a Developer Needs to Know About MCP with Neo4j