World Congress 2026 North America • Sep 25, 2026 • Session details

The New Bottleneck in Software Development

Aditya Jayaprakash

The KV cache is the new memory bottleneck crippling generative AI applications. Learn how modern architectures like Grouped-Query Attention overcome this limit to optimize your production workloads.

The New Bottleneck in Software Development thumbnail

Checking access…

Playback and chapters load privately for Free videos.

Matching moments

1:05 min

Evaluating language models by understanding underlying architectural changes

Jofia Jose Prakash Jofia Jose Prakash · World Congress 2026 North America

1:17 min

Understanding the unique architecture of LLM applications

Saloni Garg Saloni Garg · World Congress 2026 North America

2:37 min

Understanding core parameters and mechanics of large language models

Julián Duque Julián Duque · World Congress 2025

7:28 min

Accelerating product features using generative large language models

David Singleton David Singleton +1 · Coffee With Developers

5:01 min

Leveraging large language models for code optimization and development

Stephan Gillich Stephan Gillich +3 · World Congress 2024

2:26 min

Capabilities and applications of large language models

Aditi Godbole · LIVE