World Congress 2024

Supercharge Inferencing of GenAI & LLM on AI PC

July 18, 2024 16:00 – 18:00 · 120 min Workshop room M4 (40)

What this session covers

In this session, we will introduce how to run local, fast AI inferencing for LLM and gen AI on AI PC with the help of OpenVINO, of which the main points are:

a) AI PC quick overview

b) OpenVINO Notebooks intro & demo (quick demos/GIFs of different use cases on CPU, GPU and NPU)

c) Challenges of AI deployment and general overview of OpenVINO (developer journey)

d) Introduction of main features intro, including - Model converter - NNCF (weights compression for LLM) - Optimum-Intel (GenAI) - Deployment (AUTO plugin)

Related talks at this congress

Open session

World Congress 2024

July 18, 2024 · 17:30–18:00

STAGE 3 (120)

Bringing the power of AI to your application.

Krzysztof Cieślak

Developer Tools Researcher, OSS contributor, FP enthusiast

Krzysztof Cieślak
Open session

World Congress 2024

July 18, 2024 · 10:50–11:20

STAGE 9 (600)

Supercharge your cloud-native applications with Generative AI

Cedric Clyburn

Cedric Clyburn is Developer Advocate at Red Hat

Cedric Clyburn
Open session

World Congress 2024

July 19, 2024 · 16:20–16:50

STAGE 5 (500)

Generative AI power on the web: making web apps smarter with WebGPU and WebNN

Christian Liebel

Web developer

Christian Liebel
Open session

World Congress 2024

July 19, 2024 · 12:00–13:00

Workshop room M5 (45)

Supercharge DevOps workflows with AI

SKi

Senior Solutions Architect, GitHub

SKi
All sessions at this congress