About This Session
Building a Voice AI demo takes a weekend; production engineering is a different beast. This session delivers the architectural playbook for scaling resilient, enterprise-grade voice agents. We will evaluate speech-to-speech versus modular (STT → LLM → TTS) pipelines, WebRTC stacks, and hybrid topographies for strict data residency. You will learn fast-brain/slow-brain UX patterns to keep turn-taking responsive while backends process tool calls, alongside strategies for runtime model switching and custom vocabulary. Walk away with the exact trade-offs, patterns, and metrics needed to operate low-latency voice pipelines at scale.
Topics
- Agents