21 Experiments in Six Weeks: A Playbook for Improving Your AI Agent
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Think upgrading to a smarter model automatically improves your AI agent? Sentry ran 21 A/B tests in six weeks and discovered that more expensive models can actually hurt performance.
Checking access…
Playback and chapters load privately for Free videos.
Matching moments
More from World Congress 2026 North America
Related videos
Related articles
TP
Thomas Pamminger
AH
Aishi Huang
CH
Chris Heilmann
DC
Daniel Cranney
CH
Chris Heilmann
KP
Kamen Petroff
CH
Chris Heilmann
From learning to earning
Jobs that call for the skills explored in this talk.
about 2 months ago
•
Verified
Senior AI/ML Engineer
PagerDuty
Lisbon, Portugal
Expert
Remote
AI Frameworks
AI-assisted coding tools
26 days ago
Senior AI Developer
PwC
United States
Expert
Remote
Docker
Github
Fastapi
about 2 months ago
Staff Software Engineer, GitHub Intelligence (Copilot Agents)
GitHub
San Francisco, CA, United States
Experienced
Ruby
Golang
Github
2 months ago
Software Engineer - Video
Twilio
Austin, TX, United States
Experienced
$138k–203k
Remote
NoSQL
Nagios
Twilio
20 days ago
•
Verified
Principal AI Forward Deployed Engineer
Expedia
San Jose, United States
Expert
Remote
AI Frameworks
Software Architecture
20 days ago
•
Verified
Principal Product Manager, Agentic Evals
Expedia
San Jose, United States
Expert
Remote
Product Management