World Congress 2026 Europe

Smaller Voice Models

July 9, 2026 13:25 – 13:30 · 5 min Airstream 1

What this session covers

In this talk, we’ll cover why some tasks only need smaller models: benefiting from zero cost, privacy, low latency, and offline reliability. Using voice AI as the example, we’ll explore how compact text-to-speech models can run on-device or on-premise, avoiding the cost and latency of cloud-only systems. We’ll share practical lessons from building smaller voice models, where they work well, where larger models are still useful, and how developers can think about model size as a product decision rather than just a benchmark. The session is aimed at anyone building real-time AI applications, voice agents, embedded tools, or privacy-sensitive user experiences.

Related talks at this congress

Open session

World Congress 2026 Europe

July 10, 2026 · 12:20–12:50

Stage 12

Future of Mobile AI. What On-Device Intelligence Means for App Developers

Sasha Denisov

CTO and Co-Founder of Brainform.ai

Sasha Denisov
Open session

World Congress 2026 Europe

July 10, 2026 · 16:20–16:50

Stage 6 - powered by Microsoft

Fine-Tuning Small Language Models for Agentic AI

Björn Buchhold

Technology Evangelist at CID

Björn Buchhold
Open session

World Congress 2026 Europe

July 10, 2026 · 12:15–14:15

Room M2 (40 Seats)

Compress, Cut, and Distill: The Latest Gen AI Model Compression Techniques in Practice

Sergio Perez

Senior Solution Architect at NVIDIA

Sergio Perez
Open session

World Congress 2026 Europe

July 9, 2026 · 10:10–10:40

Stage 8 - powered by Red Hat

Owning the Inference Layer: When and How to Run your Own Models

Taylor Jordan Smith

Senior AI Developer Advocate at Red Hat

Taylor Jordan Smith
All sessions at this congress