World Congress 2025 Aug 20, 2025 Session details

Prompt API & WebNN: The AI Revolution Right in Your Browser

Christian Liebel

Tired of high cloud costs and API rate limits? Discover how WebNN and the Prompt API unlock fast, offline-capable, privacy-first AI directly in the browser.

Pause
Mute Enter Fullscreen
#1 about 3 min

Drawbacks of cloud dependencies and local inference benefits

Highlighting privacy, offline capabilities, capacity guarantees, and cost reductions driving the shift toward browser-based artificial intelligence.

#2 about 2 min

Standardizing inference with working and community group APIs

Comparing the bring-your-own-model approach proposed by the webml working group against experimental built-in APIs from the community group.

#3 about 7 min

Executing open weight large language models with WebLLM

Utilizing webgpu via webllm and transformers.js to load and execute models entirely locally for text and visual tasks.

#4 about 5 min

Accelerating local neural networks using the WebNN API

Accessing neural processing units to drastically improve inference frame rates for workloads like real-time image classification.

#5 about 3 min

Solving cross-origin model size and native performance limitations

Proposing a shared cross-origin storage architecture and introducing built-in browser algorithms to mitigate repetitive gigabyte-scale model downloads.

#6 about 6 min

Implementing browser native capabilities with the Prompt API

Calling locally installed models directly through native summarizer, translation, and prompt object interfaces contextually.

#7 about 7 min

Integrating generative workflows for data classification and extraction

Moving beyond conversational interfaces by leveraging models for automated form filling, structural parsing, and real-time multimodal voice interaction.

#8 about 2 min

Tradeoffs of transitioning to on-device machine learning architectures

Weighing the strict privacy and low latency benefits against system requirements, inference speed limits, and reduced computational capability.

Matching moments

2:54 min

Leveraging Chrome AI and nano models for web applications

Raymond Camden · Perf + AI

2:46 min

The case for native AI in web browsers

Maxim Salnikov Maxim Salnikov · WWC 2025

2:17 min

Accelerating local machine learning models via WebNN

Christian Liebel Christian Liebel · WWC Europe 2026

2:46 min

Enhancing native browser experiences with on-device generative AI

Chris Heilmann +2 · LIVE

1:04 min

Running local large language models securely using paired web GPUs

Önder Ceylan Önder Ceylan · WWC 2024

51 sec

Adopting hybrid approaches for browser-based AI models

Jason Mayes · Coffee With Developers

Upcoming sessions on this topic

Open session

World Congress 2026 North America

Small LLM in your Browser: Huge Opportunities for Web Applications

Daniel Ostrovsky

UI/UX Architect at Payoneer | AI Architect | Full Cycle Development Expert | Public Speaker | Open Source Contributor |

Daniel Ostrovsky
Open session

World Congress 2026 North America

Building Stuff with GenAI - The Open Minded Workshop beyond OpenAI

Andreas Erben

CTO for Applied AI and Metaverse at daenet

Andreas Erben
Open session

World Congress 2026 North America

Agents That Own Their Inference: Building Production AI Agents on Dedicated GPUs

Duan Lightfoot

Sr. AI Engineer, Akamai

Duan Lightfoot
Open session

World Congress 2026 North America

No Single Model to Rule Them All: Building Resilient AI Agents Across Open & Closed LLMs

Emmanuel Acheampong

Senior Manager Developer Relations at Crusoe AI

Emmanuel Acheampong
Open session

World Congress 2026 North America

SecurePrompt: Building a Pre-Flight Security Layer for Agentic AI

Ravi Sastry Kadali

AI/ML Engineer at General Motors

Ravi Sastry Kadali
Open session

World Congress 2026 North America

Making Science Larger, not just Faster

Yuval Dvir

Commercial Executive, SandboxAQ

Yuval Dvir