> Markdown version of [/jobs/ext/164006-senior-product-manager-software-developer-platform](https://www.wearedevelopers.com/jobs/ext/164006-senior-product-manager-software-developer-platform). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Product Manager, Software & Developer Platform - **Company:** quadric, Inc - **Location:** Burlingame, CA, United States - **Experience:** Expert - **Salary:** $200,000.0 - $250,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Computing Platforms, Cursor (Graphical User Interface Elements), Release Management, Software Engineering, Large Language Models, ONNX (Open Neural Network Exchange) Format, TensorRT - **Published:** May 16, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=cca78a9f3034806c ## About the Role Do you have experience in Release management?, * Domain - non-negotiable. Shipped product on at least one of: NPU or AI accelerator IP/silicon stack; graph or ML compiler (TVM, MLIR, XLA, or proprietary); developer-facing AI inference runtime or agent framework. "Adjacent" does not count. * Modern AI workload fluency - non-negotiable. Ready conversation, no prep, on: agentic workflows and LLM serving, KV cache optimization, quantization schemes (AWQ, GPTQ, SmoothQuant, QAT vs. PTQ), datatypes (INT4, FP8, BF16, OCP MX), and inference platforms (vLLM, llama.cpp, TensorRT-LLM, ExecuTorch, ORT). * Shipping bar - non-negotiable. You shipped a developer-facing AI or compute product - SDK, runtime, compiler, or inference service - with real users and a release cadence you owned. * Agent-pilled - non-negotiable. You use agentic AI tools daily (Claude Code, Cursor, or equivalent) to produce work. Having read about agentic AI without integrating it is not sufficient. * Customer first - non-negotiable. When engineering wants to build the elegant thing and the customer needs the workable thing, you take the workable thing every time. * 5+ years in PM, with 3+ years on a developer-facing AI/ML or compute platform * Owned a release cadence - picked what ships, what slips, and defended the call * Experience with in-person technical customer reviews * Bay Area resident or willing to relocate to Burlingame Preferred * ML background: graduate degree, published work, trained model, or OSS contribution * Automotive Tier 1 engagement; ISO 26262 awareness * Prior product work on a competing NPU/GPU/AI accelerator stack (MetaWare, Arm ML, Ceva, TensorRT, Hailo, Tenstorrent, etc.) * OSS contributions to vLLM, llama.cpp, TVM, MLIR, ONNX Runtime, or ExecuTorch ## Description * Software release train. Own the monthly SDK release and quarterly major: contents, release sync, go/no-go, release notes, and customer communications. * Pattern coverage roadmap. Decide which graph patterns CGC compiles next - attention variants, quantization schemes, normalization patterns - and sequence them against customer model requirements each quarter. * Market-driven demo strategy. Lead with the market story and push it through every layer: demo, model zoo, pattern coverage, compiler work. Own what we publish and when. * Customer engagement. Present in technical reviews with anchor customers. Convert model gap lists into engineering-ready roadmap entries. * Quantization and numerics. Own the roadmap for INT4 (W4A8/W4A16), FP8, OCP MX, and KV cache compression. Coordinate with HW PM on MAC capability and with customer SW teams on model format decisions ahead of silicon tape-out. * Framework and runtime integrations. Define the integration strategy for GGML/llama.cpp, vLLM, ONNX Runtime, ExecuTorch, and HF Optimum - deep partnership vs. thin reference. * Model zoo. Maintain a set of customer-confidence models (LLM chat, BEVFormer, VLA, ADAS perception) that serve as forcing functions for compiler completeness. * Quarterly roadmap tours. Take the roadmap to anchor customers, prospects, and the field. Brief the PMM monthly on what shipped and how to position it. * Competitive intelligence. Track Synopsys MetaWare, Arm KleidiAI, Ceva NeuPro Studio, and NVIDIA TensorRT-LLM. Brief exec and sales quarterly. * Safety and quality. Coordinate with the safety lead on ISO 26262 traceability and qualification artifacts. ## Related Videos - [Modern Web Development with Nuxt3](https://www.wearedevelopers.com/videos/294-modern-web-development-with-nuxt3) - [How to govern Vibe Coding for the Enterprise](https://www.wearedevelopers.com/videos/100290-how-to-govern-vibe-coding-for-the-enterprise) - [Tour de Force: Open-Source LLM Inference Optimization from Simple to Sophisticated](https://www.wearedevelopers.com/videos/100099-tour-de-force-open-source-llm-inference-optimization-from-simple-to-sophisticated) - [Localized Open Models in Production: What Builders Need to Know](https://www.wearedevelopers.com/videos/100270-localized-open-models-in-production-what-builders-need-to-know) - [Architecting the Future: Leveraging AI, Cloud, and Data for Business Success](https://www.wearedevelopers.com/videos/1096-architecting-the-future-leveraging-ai-cloud-and-data-for-business-success) - [Efficient deployment and inference of GPU-accelerated LLMs​](https://www.wearedevelopers.com/videos/929-efficient-deployment-and-inference-of-gpu-accelerated-llms) ## Related Articles - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it)