Senior Product Manager - AI Platform Inference

NVIDIA Ltd.
Santa Clara, CA, United States
6 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Working hours
Regular working hours

Tech stack

Computer Engineering Machine Learning Performance Tuning Software Product Management Software Deployment Software Engineering Graphics Processing Unit (GPU) Large Language Models AI Platforms Information Technology TensorRT

Job description

Experteer Overview In this role you drive the product vision and execution for AI platform inference on NVIDIA GPUs. You collaborate with developers across the ecosystem to accelerate deployments and optimize performance. You shape roadmaps, strategy, and go-to-market plans to help developers build better inference deployments at scale. Your work enables NVIDIA to lead in AI deployment tooling and open opportunities for innovation and impact in GenAI workloads. Compensation / Benefits * Create products that help developers build better Inference deployments * Develop product strategy, roadmaps, and go-to-market plans * Collaborate with internal and external developers to build roadmaps for model optimization software * Align with leadership to drive the company strategy Tasks * Experience with Inference deployment and optimization software (e.g., vLLM, SGLang, FlashInfer, TensorRT-LLM, Triton, Dynamo, TorchAO) * Demonstrable knowledge of GenAI or ML concepts, particularly performance optimization and software development/delivery * BS or MS in Computer Science, Computer Engineering, or equivalent experience * 6+ years of technical product management experience at a technology company * Strong communication and interpersonal skills Key requirements * equity * benefits

Requirements

Experteer Overview In this role you drive the product vision and execution for AI platform inference on NVIDIA GPUs. You collaborate with developers across the ecosystem to accelerate deployments and optimize performance. You shape roadmaps, strategy, and go-to-market plans to help developers build better inference deployments at scale. Your work enables NVIDIA to lead in AI deployment tooling and open opportunities for innovation and impact in GenAI workloads. Compensation / Benefits * Create products that help developers build better Inference deployments * Develop product strategy, roadmaps, and go-to-market plans * Collaborate with internal and external developers to build roadmaps for model optimization software * Align with leadership to drive the company strategy Tasks * Experience with Inference deployment and optimization software (e.g., vLLM, SGLang, FlashInfer, TensorRT-LLM, Triton, Dynamo, TorchAO) * Demonstrable knowledge of GenAI or ML concepts, particularly performance optimization and software development/delivery * BS or MS in Computer Science, Computer Engineering, or equivalent experience * 6+ years of technical product management experience at a technology company * Strong communication and interpersonal skills Key requirements * equity * benefits

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · WWC Europe 2026

2:36 min

Applying supervised machine learning for practical rule extraction

Katja Träumner

2:36 min

Choosing between managed AI platforms and custom governance

Péter Farkas Péter Farkas · Europe 2026 Virtual

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · WWC 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:32 min

Core libraries driving inference engines and multi-GPU networking

Adolf Hohl Adolf Hohl · WWC 2024

Videos

See all

Related articles

See all