World Congress 2026 North America

Finding the Edges: Testing, Evaluating, and Monitoring Voice AI Agents Before Your Users Do

September 23–25, 2026

World Congress 2026 North America

September 23–25, 2026 · San José, CA

Attend in person

Get tickets

Watch remotely

Watch live with Pro

Pro

Can’t make it to San José? Watch this session live with Pro. You also get:

  • All full videos, bookmarks, and playlists
  • World Congress livestreams
See pricing

What this session covers

Every demo of a voice AI agent looks great — until real users start talking. They interrupt, mumble, switch languages mid-sentence, and ask the one question that sends the agent off the rails. Traditional scripted QA only verifies the behaviors you already thought of, which is exactly why so many voice agents fail in production in ways their teams never saw coming.

This session walks through a practical, engineering-grade approach to shipping Voice AI agents you can trust, built on three pillars: simulation, evaluation, and monitoring. You’ll see how synthetic “drivers” — AI-powered simulated users with distinct personalities, goals, and contexts — hold realistic multi-turn conversations with your agent across 30+ languages and real-world audio conditions (noise, crosstalk, clipping), actively exploring the edges scripted tests miss. We’ll then look at how judge-based, symbolic, and audio evaluations turn those discoveries into CI/CD release gates, so a change that degrades conversation quality fails the build before it reaches users. Finally, we’ll close the loop with production monitoring that captures real failures and automatically converts them into regression tests — so the same mistake never ships twice.

You’ll leave with a concrete blueprint for finding the edges of your Voice AI agent before your customers do.

Related talks at this congress

Open session

World Congress 2026 North America

Your Evals Passed. Your Agent Just Emptied a Database.

Tejas Pravinbhai Patel

IEEE Award-Winning Researcher | Best Keynote Speaker | Sr. Software Engineer at Amazon | AI Systems & Agent Architect

Tejas Pravinbhai Patel
Open session

World Congress 2026 North America

Taming Rogue Agents: Observability-Driven Evaluation for Production Reliability

Anagha Rumade, Anjana Umapathy, Apoorva Jaiswal

Anagha Rumade
Anjana Umapathy
Apoorva Jaiswal
Open session

World Congress 2026 North America

Closing the Visibility Gap: Lessons from Safety Critical Agentic Systems

Vivek Pandit

Principal Engineer at Cadence

Vivek Pandit
Open session

World Congress 2026 North America

I Don't Trust AI Agents (And Neither Should You): Building Production-Ready Architectures

Darko Mesaros

Distinguished Developer Advocate at AWS

Darko Mesaros
All sessions at this congress