World Congress 2026 North America
September 24, 2026 · 14:10–14:40
Stage 1
Anatomy of an AI Request: Where Latency and Cost Are Really Born
Dan Fu
VP of Kernels at Together AI
World Congress 2026 North America
World Congress 2026 North America
September 23–25, 2026 · San José, CA
Attend in person
Get ticketsWatch remotely
Pro
Can’t make it to San José? Watch this session live with Pro. You also get:
In the rapidly evolving landscape of computing, Graphics Processing Units (GPUs), Language Processing Units (LPUs), and Tensor Processing Units (TPUs) play pivotal roles in accelerating complex tasks, particularly in machine learning and artificial intelligence.
GPUs are renowned for their parallel processing capabilities, making them ideal for rendering graphics and handling large datasets. LPUs are specialized for optimizing natural language processing tasks, enhancing efficiency in understanding and generating human language. TPUs, developed by Google, are tailored specifically for training and inference of machine learning models, offering significant performance advantages for large-scale AI applications.
As we explore these technologies, we’ll also look at emerging processing units designed for specific AI use-cases and the future of computational advancements.
Join me to dive into the intricacies of these processing units, their applications, and what lies ahead in the world of computing technology.
World Congress 2026 North America
September 24, 2026 · 14:10–14:40
Stage 1
Dan Fu
VP of Kernels at Together AI
World Congress 2026 North America
September 24, 2026 · 14:10–14:40
Stage 7
Wayne Liu
Chief Growth Officer and Americas President of Perfect Corp.
World Congress 2026 North America
September 23, 2026 · 10:45–12:45
Stage 10
Duan Lightfoot
Sr. AI Engineer, Akamai
World Congress 2026 North America
September 24, 2026 · 14:10–14:40
Stage 5
Nitin Eusebius
AWS - Principal Solutions Architect