World Congress 2026 North America • Sep 24, 2026 • Session details

Taming Rogue Agents: Observability-Driven Evaluation for Production Reliability

Anagha Rumade , Anjana Umapathy , Apoorva Jaiswal

Rogue AI agents have already issued unauthorized refunds and sold cars for a dollar. Discover how to prevent production disasters by securing reasoning chains with observability-driven evaluation.

Taming Rogue Agents: Observability-Driven Evaluation for Production Reliability thumbnail

Checking access…

Playback and chapters load privately for Free videos.

Matching moments

2:40 min

Reviewing live performance of self-correcting AI engineering agents

Ingo Eichhorst Ingo Eichhorst · World Congress 2026 Europe

1:33 min

Identifying major gaps in production AI agents

Ashok Prakash Ashok Prakash · World Congress 2026 North America

1:56 min

Introduction to evaluating AI code review agents

Sofia Rest Sofia Rest · World Congress 2026 North America

2:28 min

Introduction to building reliable AI agents in production

Max Tkacz Max Tkacz · World Congress 2025

2:43 min

Best practices for implementing reliable AI agent frameworks

Marcel Scherenberg Marcel Scherenberg · World Congress 2025

3:34 min

Architecting autonomous agents for production lifecycle management

Sebastian Kister Sebastian Kister · World Congress 2026 Europe