World Congress 2025

Scrape, Train, Predict: The Lifecycle of Data for AI Applications

July 11, 2025 15:00 – 15:30 · 30 min Stage 8

What this session covers

Behind every AI breakthrough lies a critical dependency: fresh, high-quality web data. This session explores the symbiotic relationship between web scraping and artificial intelligence that’s transforming how developers build data-driven applications. Discover how machine learning is solving web scraping’s biggest challenges - from bypassing sophisticated anti-bot measures with ML-powered response recognition to building adaptive parsers that automatically handle website structure changes. We’ll explore practical use cases: training AI models with scraped data, powering real-time generation systems, and building resilient data pipelines that don’t break when websites update. You’ll learn how AI transforms web scraping from a maintenance nightmare into a strategic advantage, whether you’re building recommendation engines, training custom models, or creating intelligent data collection systems.

Related talks at this congress

Open session

World Congress 2025

July 10, 2025 · 13:30–14:00

Stage 10

How to scrape modern websites to feed AI agents

Jan Curn

Founder and CEO of Apify

Jan Curn
Open session

World Congress 2025

July 9, 2025 · 11:00–13:00

M3 (45 Seats)

AI and LLM in .NET: Transforming Big Data Analysis

Mehdi

Halocline

Mehdi
Open session

World Congress 2025

July 11, 2025 · 09:30–11:30

M4 (40 Seats)

Learn to Build Agentic AI Workflows for Enterprise Applications

Lavinia Ghita, Ziv Ilan

Lavinia Ghita
Ziv Ilan
Open session

World Congress 2025

July 10, 2025 · 10:45–12:45

M6 (40 Seats)

Beyond AI Readiness: Scaling Data-Driven Engineering for Next-Gen Software & Infrastructure

Eduard, Jannis Eickenroth, Sophie Kleinke

Eduard
Jannis Eickenroth
Sophie Kleinke
All sessions at this congress