World Congress 2025

Evaluating AI models for code comprehension

July 11, 2025 09:00 – 09:30 · 30 min Stage 11
ai llms productivity performance devops nlp

What this session covers

With so many AI models out there, which one is actually the best at understanding code?

As we’ve built Graphite Reviewer, our high-signal AI-powered code review companion, we’ve run thousands of tests, gotten tons of feedback from users, and have extensively examined how these various models perform in real world scenarios.

In this talk we’ll explore the nuances between these models, how they perform specifically in relation to code comprehension, and finally answer the question: which AI model should developers be using?

This talk will include a deep dive into the various experiments we’ve run on AI code comprehension, comparing all of the top AI models today, and an exploration of the results and data we’ve collected.

Related talks at this congress

Open session

World Congress 2025

July 11, 2025 · 16:20–16:50

Stage 2

How we built an AI-powered code reviewer in 80 hours

Yan Cui

Developer Advocate at Lumigo

Yan Cui
Open session

World Congress 2025

July 11, 2025 · 09:00–09:30

Stage 7

AI-Powered Code Documentation: Simplify the Complex

Patrick Schnell

Patrick Schnell, AI-enabler and CEO at schnell.digital GmbH

Patrick Schnell
Open session

World Congress 2025

July 10, 2025 · 13:25–13:30

Stage 8

Leapter: The Reinvention of Software Development? A Future Built On AI Generated Code.

Robert Werner

Leapter CTO and Founder

Robert Werner
Open session

World Congress 2025

July 10, 2025 · 15:30–16:00

Stage 3 - Microsoft

Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models

Kevin Lewis, Sandra Ahlgrimm

Kevin Lewis
Sandra Ahlgrimm
All sessions at this congress