> Markdown version of [/events/world-congress-2025/sessions/538-a-deep-dive-on-how](https://www.wearedevelopers.com/events/world-congress-2025/sessions/538-a-deep-dive-on-how). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # A Deep Dive on How To Leverage the NVIDIA GB200 for Ultra-Fast Training and Inference on Kubernetes - **Date:** Thursday, Jul 10, 2025 - **Time:** 11:30–12:00 (30 min) - **Room:** Stage 7 - **Event:** World Congress 2025 - **Tags:** ai, cloud, kubernetes ## Recording [Watch recording](https://www.wearedevelopers.com/videos/1625-a-deep-dive-on-how-to-leverage-the-nvidia-gb200-for-ultra-fast-training-and-inference-on-kubernetes) ## Description Kubernetes traditionally does not have a mechanism for allocating non-node-local resources. The one exception being persistent volumes, which allow a user to attach the same volume to multiple pods running on different nodes. With the introduction of Dynamic Resource Allocation (DRA) we now have a way to allocate any type of resource with similar semantics. In this talk, we discuss how DRA’s ability to allocate non-node-local resources has unlocked the potential to read / write remote GPU memory over high-bandwidth, multi-node NVLinks. We begin with an introduction on how DRA models non-node-local resources in general, followed by the specifics of how we have leveraged this capability to enable lightning fast multi-node training and inference on the NVIDIA GB200 NVL72 supercomputer. As part of this, we discuss how this support has been pushed to all major cloud providers and integrated with their managed Kubernetes offerings. We conclude with a demo. ## Speaker ### [Kevin Klues](https://www.wearedevelopers.com/@kevin-klues) Distinguished Engineer at NVIDIA ## Related talks at this congress - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/events/world-congress-2025/sessions/586-your-next-ai-needs) — Anshul Jindal, Martin Piercy - [NVIDIA Expert Session: Your AI, Everywhere: Unifying Infrastructure for Unbounded Innovation](https://www.wearedevelopers.com/events/world-congress-2025/sessions/839-nvidia-expert) — Anshul Jindal, Martin Piercy - [NVIDIA Expert Session: Optimized Deployment of Large Language Models](https://www.wearedevelopers.com/events/world-congress-2025/sessions/491-nvidia-expert) — Anshul Jindal, Kevin Klues, Martin Piercy - [NVIDIA Expert Session: Accelerating Data Analytics Workflows on GPU Systems](https://www.wearedevelopers.com/events/world-congress-2025/sessions/753-nvidia-expert) — Miguel Martínez, Roman