> Markdown version of [/jobs/ext/1501016-manager-distinguished-engineer-dgx-systems](https://www.wearedevelopers.com/jobs/ext/1501016-manager-distinguished-engineer-dgx-systems). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Manager, Distinguished Engineer - DGX Systems... - **Company:** NVIDIA Ltd. - **Location:** Santa Clara, CA, United States - **Experience:** Expert - **Salary:** $320,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, BIOS, Nvidia CUDA, Data Centers, Firmware, PCI Express, Performance Tuning, Software Engineering, Systems Architecture, System Testing, Systems Integration, Network Switches, Information Technology, Hardware Infrastructure, Nim (Programming Language), Microservices - **Published:** July 30, 2026 - **Apply:** https://www.juju.com/job/00000000gkh8ot ## About the Role + BS or MS in Computer Science, Electrical Engineering, or related field or equivalent experience. + 12+ overall years in systems firmware/software engineering, with 5+ years in engineering leadership. + Deep expertise in server system stack including SBIOS, BMC, OS, applications and system-level integration of complex multi-component products. + Proven track record delivering multi-generation server or data center platforms from architecture through customer deployment. + Experience managing engineering organizations across multiple geographies in a matrix environment. + Strong understanding of server hardware: CPU, GPU, interconnect, memory, PCIe, power delivery. + Experience owning end-to-end product quality-from firmware validation through full-stack system testing to field deployment. Ways to stand out from the crowd: + Experience with NVIDIA DGX, or GPU-accelerated server platforms.Track record driving server bring-up for new silicon and system architecture redesigns. + Familiarity with DMTF Redfish, OCP standards, and server manageability ecosystems. + Experience with AI/DL workload validation and performance optimization at the platform level.Demonstrated ability to operate at VP/SVP level, influencing cross-BU strategic decisions. ## Description We are seeking an engineering leader responsible for end-to-end delivery of every DGX compute system-from firmware through the AI stack to customer deployment. You will ensure each DGX product ships as a production-ready system where firmware, OS, drivers, CUDA, networking, and AI applications work together seamlessly, while driving architecture and roadmap for next-generation platforms. What you'll be doing: + End-to-End Stack Readiness: Ensure every DGX platform is ready for the full NVIDIA software stack-firmware, DGX OS, GPU drivers, CUDA toolkit, DCGM, DOCA/OFED, and management tools-as a validated, production-quality product. Own the GA SW/FW release process delivering firmware bundles, BaseOS ISOs, and release notes to OEM/OSV partners. Ensure platforms support AI agents like NemoClaw, Hermes agents, NIM microservices, and workloads customers expect out of the box. + Platform Firmware Development: Lead development of the manageability firmware stack (BMC, BIOS) for all DGX platforms. Ensure firmware from partner teams (GPU, CPU, networking) integrates correctly at system level. Manage 3rd-party vendors and drive platform requirements (NVPOR) across all firmware areas. + Validation Strategy: Define validation strategy proving each DGX platform is production-ready: end-to-end system validation including firmware regression, NVQual certification, DL workload performance, OS/CUDA stack testing, multi-user scenarios, power/thermal validation, and field upgrade reliability. Establish quality gates and zero ship-stopper discipline. + Platform Bring-Up & Architecture: Drive platform bring-up for each new DGX system-coordinating first boot across new silicon (CPU, GPU), board design, and firmware teams. Own architectural strategy for next-generation platforms including firmware update mechanisms, system security posture, and AI application readiness. + Customer Deployment & Enablement: Ensure firmware release flows meet CSP and enterprise deployment requirements. Represent DGX platform readiness in executive reviews and strategic planning with VP/SVP leadership. Engage with industry standards bodies (DMTF Redfish, OCP). + Product Delivery Lifecycle: Own the complete DGX delivery lifecycle-system architecture, firmware development, integration, full-stack validation, GA release, and customer deployment-for every DGX product. + Cross-Org Alignment: Serve as single point of accountability for DGX platform readiness across NVIDIA-aligning GPU, CPU, networking, security, OS, and AI software teams to deliver on schedule. + Quality & Vendor Management: Own RCCA processes for field issues. Manage external vendor partnerships (AMI for SBIOS, BMC contributors) with clear quality gates and program tracking. + Team Leadership: Build and lead a world-class engineering organization. Mentor and develop leaders. Foster a culture of technical excellence, intellectual honesty, and customer obsession. ## Related Videos - [Coffee with Developers - Stephen Jones - NVIDIA](https://www.wearedevelopers.com/videos/1303-coffee-with-developers-stephen-jones-nvidia) - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/videos/1590-your-next-ai-needs-10-000-gpus-now-what) - [10M Data Records Lost, Underwater Computing, and Psychedelic Fish - Matthias Geniar](https://www.wearedevelopers.com/videos/1908-10m-data-records-lost-underwater-computing-and-psychedelic-fish-matthias-geniar) - [Playing Pong on a shoulder press machine](https://www.wearedevelopers.com/videos/100140-playing-pong-on-a-shoulder-press-machine) - [A Deep Dive on How To Leverage the NVIDIA GB200 for Ultra-Fast Training and Inference on Kubernetes](https://www.wearedevelopers.com/videos/1625-a-deep-dive-on-how-to-leverage-the-nvidia-gb200-for-ultra-fast-training-and-inference-on-kubernetes) - [Building the Nervous System of AI - Michael Kagan (NVIDIA)](https://www.wearedevelopers.com/videos/2133-building-the-nervous-system-of-ai-michael-kagan-nvidia) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere)