World Congress 2025 Aug 20, 2025 Session details

Hello JARVIS - Building Voice Interfaces for Your LLMS

Nathaniel Okenwa

Why do most AI voice bots feel awkwardly robotic? Learn to build intuitive, JARVIS-like voice interfaces by engineering responsive state routing, fluid interruptions, and composable LLM pipelines.

Pause
Mute Enter Fullscreen
#1 about 2 min

Introduction to real-time communication programming interfaces

How modern application programming infrastructures facilitate voiceover real-time messaging pipelines.

#2 about 4 min

Designing voice interfaces inspired by fictional artificial intelligence

Why science fiction uses ambient computing and vocal communication as intuitive shortcuts for human-computer interaction.

#3 about 2 min

Enhancing conversational intent through modern large language models

How unscripted interactions and action-oriented tool paths resolve the contextual brittleness of early virtual assistants.

#4 about 5 min

Navigating the conversational uncanny valley in voice systems

Managing psychological reactions to machine pacing and non-verbal gaps to ensure flowing interaction workflows.

#5 about 2 min

Architecting a web real-time communication stack for agents

Routing live audio streams through decoupled transcription and generative endpoints to facilitate smooth conversations.

#6 about 3 min

Evaluating decoupled conversational architectures versus end-to-end models

Why modular sequence pipelines offer better logic injection and behavioral fine-tuning than monolithic processing environments.

#7 about 8 min

Programming contextual interruption logic for responsive machine interaction

Tracking active playback states to dynamically pause out-of-date audio outputs and append localized interruption history.

#8 about 5 min

Implementing voice interstitials to mask background computational latency

Evaluating background network tasks to emit temporary auditory dialog while complex external data sources resolve.

#9 about 3 min

Orchestrating managed communication platforms for enterprise production readiness

Offloading intensive timing states and concurrent socket delivery back to centralized conversational cloud relay services.

Matching moments

2:05 min

Challenges of managing voice architecture and latency

Chris Heilmann +3 · LIVE

2:11 min

Analyzing voice interface research projects and technical limitations

Tobias Münch Tobias Münch · WWC 2024

4:16 min

Addressing latency and architecture in voice agents

Chris Heilmann +2 · LIVE

1:44 min

Summarizing developer experience and artificial intelligence companions

Robert Hoffmann Robert Hoffmann +1 · LIVE

2:44 min

Understanding language models and autonomous executing agents

Chris Heilmann +2 · LIVE

1:29 min

Designing accessible voice content and interfaces

Ana Rodrigues · LIVE

Upcoming sessions on this topic

Open session

World Congress 2026 North America

From Software Agents to Physical Devices: Inside the Agentic Hardware Stack

Michael Yuan, Vivian Hu

Michael Yuan
Vivian Hu
Open session

World Congress 2026 North America

AI Agents are Only as Smart as their Context: Building a Real-Time Context Engine at Intuit

Bharat Patel

Lead Software Engineer at Intuit

Bharat Patel
Open session

World Congress 2026 North America

Proactive AI That Doesn’t Annoy Users: Building Context-Aware Notification Systems

Raju Dandigam Dandigam

Engineering Manager at Navan

Raju Dandigam Dandigam
Open session

World Congress 2026 North America

No Single Model to Rule Them All: Building Resilient AI Agents Across Open & Closed LLMs

Emmanuel Acheampong

Senior Manager Developer Relations at Crusoe AI

Emmanuel Acheampong
Open session

World Congress 2026 North America

Closing the Visibility Gap: Lessons from Safety Critical Agentic Systems

Vivek Pandit

Principal Engineer at Cadence

Vivek Pandit
Open session

World Congress 2026 North America

Context Engineering Kung Fu

Carl Lapierre

Tech Lead and AI Engineer at Osedea

Carl Lapierre