World Congress 2026 Europe

Outclassing Frontier LLMs at Extracting Information

July 10, 2026 14:20 – 14:50 · 30 min Stage 1

What this session covers

Accurately extracting information from documents has been a decades-old dream. Many important workflows, from automated back-office processing to enterprise RAG, depend on it. General-purpose LLMs promise to fulfill this dream, but they have drawbacks: they make mistakes, are expensive, and are difficult to use privately.

The solution: specialized LLMs, built specifically for document extraction.

In this talk, I will present NuExtract3, the leading document-extraction LLM specialized in both structured extraction (turning documents into JSON following a schema) and OCR (turning documents into clean Markdown). I will demonstrate its capabilities, discuss what makes it different from general-purpose LLMs, and show how developers can use it either through its open-source version or through the NuExtract Platform.

Related talks at this congress

Open session

World Congress 2026 Europe

July 10, 2026 · 14:20–14:50

Stage 8 - powered by Red Hat

Garbage In, Garbage Out: Engineering Reliable AI Document Extraction Pipelines

Nazeer Saeed

Staff Solutions Engineer at Apryse

Nazeer Saeed
Open session

World Congress 2026 Europe

July 10, 2026 · 09:45–11:45

Room M8 (60 Seats)

From Hallucination to Justification: Hands-On Explainability for LLMs

Lucía Conde-Moreno, Tessel Haagen

Lucía Conde-Moreno
Tessel Haagen
Open session

World Congress 2026 Europe

July 10, 2026 · 15:40–16:10

Stage 6 - powered by Microsoft

The LLM Evolution: From Sequence Imitation to Verifiable Reasoning

Kamen Petroff

Software Developer at ATOS

Kamen Petroff
Open session

World Congress 2026 Europe

July 9, 2026 · 14:50–15:20

Stage 6 - powered by Microsoft

Rules, Heuristics, or LLMs? Lessons from Solving the Same Problem Twice

Artur Naumenko

Senior Software Engineer at Softeta

Artur Naumenko
All sessions at this congress