World Congress 2025
July 11, 2025 · 16:20–16:50
Stage 3 - Microsoft
GenAI Unpacked: Beyond Basic
Damir
daenet GmbH - ACP Digital, Microsoft Regional Director, Most Valuable Professional - AI
World Congress 2025
This talk introduces the world of generative AI, focusing on Text-to-Image, Text-to-Audio, and Text-to-Video technologies for creating images, music, and videos. We explore how neural networks use diffusion models and Transformer architectures to generate outputs from short text prompts. Highlighting tools like Sora and Midjourney, we delve into techniques like Latent Diffusion Models, which combine text understanding with denoising processes to create and edit media.
A detailed look at video generation with Sora illustrates how it compresses and reconstructs visual data into final videos. We also discuss alternatives like RunwayML and SunoAI to showcase a wide range of tools for image, audio, and video generation.
By the end of this talk, you’ll gain a foundational understanding of diffusion models, an overview of generation tools, and insights into their functionality. Practical demos will provide hands-on examples throughout the presentation.
World Congress 2025
July 11, 2025 · 16:20–16:50
Stage 3 - Microsoft
Damir
daenet GmbH - ACP Digital, Microsoft Regional Director, Most Valuable Professional - AI
World Congress 2025
July 10, 2025 · 12:10–12:40
Stage 5
Maxim Salnikov
App Innovation Business Lead at Microsoft, Tech Communities Lead, Keynote Speaker
World Congress 2025
July 9, 2025 · 09:00–17:00
HIDE Room 1
Christian Weyer, Sebastian Gingter
World Congress 2025
July 10, 2025 · 13:00–15:00
M3 (45 Seats)
Michael Hall
Developer Evangelist