> Markdown version of [/@amit-kushwaha](https://www.wearedevelopers.com/@amit-kushwaha). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Amit Kushwaha Optimizes LLM inference and agentic AI deployment at NVIDIA. ## About Amit Kushwaha makes large language models run fast in production. As a Principal Solutions Architect at NVIDIA, he focuses on inference optimization and building agentic AI systems. Before this, he led AI engineering at SambaNova Systems and tackled physical machine learning problems at ExxonMobil. His work gets into the practical details of low-latency inference, breaking down methods like speculative decoding and TensorRT-LLM to serve models efficiently. Amit holds a Ph.D. in Engineering from Stanford University, where his research began in scientific computing and large-scale simulations. ## Past Sessions ### World Congress 2026 Europe · July 8, 2026 Berlin, Germany - [Faster Together: Train and Deploy a Speculative Decoding Model for Low-Latency LLM Inference](https://www.wearedevelopers.com/events/world-congress-2026-europe/sessions/1142-faster-together) · 120 min ## Links - [LinkedIn](https://uk.linkedin.com/in/amit-kushwaha28)