World Congress 2026 Europe Jul 10, 2026 Session details

AI Agents Face-Off: Same App, Multiple Frameworks

Elaine Dias Batista

Which mobile framework wins when AI writes all the code? Discover how this multi-agent showdown proves that generating functional apps requires deeply restrictive, deterministic development environments.

Pause
Mute Enter Fullscreen
#1 about 4 min

Escaping human bias when comparing mobile frameworks

Delegating identical mobile app builds to artificial agents avoids personal framework biases and yields measurable comparison data.

#2 about 5 min

Iterating from manual to automated code generation workflows

Evolving developer instructions and evaluation workflows from manual trials and basic text guidelines to formalized context files.

#3 about 3 min

Eliminating code evaluation bias and tracking execution costs

Evaluating earlier experiments reveals the anti-pattern of model self-grading and shifts measurement from token usage to distinct financial costs.

#4 about 4 min

Designing an automated system harness for continuous testing

Creating a multi-loop feedback architecture with external modules enables structured planning and fully automated interface testing validation.

#5 about 2 min

Judging code quality with anonymous language model tournaments

Establishing a ranking system where external models perform blind evaluations guarantees completely objective output code quality measurements.

#6 about 4 min

Troubleshooting rogue agent behavior and mobile automation challenges

Overcoming unexpected complications where autonomous software changed accessibility features, triggered lock screens, and spawned unauthorized emulators.

#7 about 3 min

Analyzing benchmark results for coding agents and frameworks

Performance metrics demonstrate how specific model variants create higher quality applications with fewer errors and overall lower generation costs.

#8 about 2 min

Balancing deterministic framework architecture and probabilistic code generation

Implementing strict standard boilerplate within target frameworks simplifies functionality expectations compared to probabilistic raw project creation.

#9 about 3 min

Core engineering principles for an automated testing harness

Managing dynamic generation output requires strictly deterministic structures, environment version pinning, and disciplined context prompt budgeting.

#10 about 3 min

Assembling the complete execution toolchain for local agents

Executing identical software construction tests demands extensive orchestration scripts, target dependencies, specialized profilers, and platform build utilities.

Matching moments

3:06 min

The role of artificial intelligence in mobile development

Sasha Denisov Sasha Denisov

2:34 min

Balancing developer autonomy with the adoption of coding agents

Clemens Wasner Clemens Wasner +4 · WWC Europe 2026

2:06 min

Rethinking team structures around AI agent capabilities

Mike Mike · WWC 2025

1:53 min

Validating autonomous code generation with robust automated testing

Ahmed Tikiwa Ahmed Tikiwa · WWC Europe 2026

2:28 min

Introduction to building reliable AI agents in production

Max Tkacz Max Tkacz · WWC 2025

5:27 min

Evaluating AI agents through unpredictable behavior and logic tests

Chris Heilmann +2 · LIVE

Upcoming sessions on this topic

Open session

World Congress 2026 North America

RoboCoders: Judgment Day: AI-Assisted Engineering Applied - The Battle of Agents

Viktor Gamov, Baruch Sadogursky

Viktor Gamov
Baruch Sadogursky
Open session

World Congress 2026 North America

Agents Can't Iterate Against Tests That Lie

Rocky Warren

Senior Staff Software Engineer at Clipboard

Rocky Warren
Open session

World Congress 2026 North America

No Single Model to Rule Them All: Building Resilient AI Agents Across Open & Closed LLMs

Emmanuel Acheampong

Senior Manager Developer Relations at Crusoe AI

Emmanuel Acheampong
Open session

World Congress 2026 North America

From Idea to Production with AI: Agentic Development in Practice

Daniel Ostrovsky

UI/UX Architect at Payoneer | AI Architect | Full Cycle Development Expert | Public Speaker | Open Source Contributor |

Daniel Ostrovsky
Open session

World Congress 2026 North America

DeepAgents: Build Multi-Agent AI Systems That Actually Work

Apoorva Jaiswal, Anjana Umapathy, Anagha Rumade

Apoorva Jaiswal
Anjana Umapathy
Anagha Rumade
Open session

World Congress 2026 North America

Closing the Visibility Gap: Lessons from Safety Critical Agentic Systems

Vivek Pandit

Principal Engineer at Cadence

Vivek Pandit