World Congress 2026 North America • Sep 24, 2026 • Session details

Understanding LLM Architectures: Inside the Design of Modern Models

Jofia Jose Prakash

Treating LLMs as static black boxes introduces significant production risk. Master underlying mechanics like MoE and KV caching to optimize your real-world memory latency and system throughput.

Understanding LLM Architectures: Inside the Design of Modern Models thumbnail

Checking access…

Playback and chapters load privately for Free videos.

Matching moments

1:01 min

Exploring the internal architecture of large language models

Aditya Jayaprakash Aditya Jayaprakash · World Congress 2026 North America

1:17 min

Understanding the unique architecture of LLM applications

Saloni Garg Saloni Garg · World Congress 2026 North America

1:37 min

Comparing architectural shifts across modern language models

Aditya Jayaprakash Aditya Jayaprakash · World Congress 2026 North America

2:43 min

Aligning model architectures with specific production workloads

Aditya Jayaprakash Aditya Jayaprakash · World Congress 2026 North America

2:37 min

Understanding core parameters and mechanics of large language models

Julián Duque Julián Duque · World Congress 2025

1:14 min

Exploring the architectural reasoning logic of emerging language models

Chris Heilmann Chris Heilmann +1 · LIVE