World Congress 2024

Efficient Large Language Model Customization with NVIDIA NeMo Framework.

July 18, 2024 14:50 – 15:20 · 30 min Virtual Stage 1

What this session covers

In the rapidly evolving field of artificial intelligence, efficiently customizing large language models (LLMs) is essential for high performance across diverse applications. This talk will cover advanced techniques for parameter-efficient fine-tuning (PEFT) of LLMs, including p-tuning, prompt-tuning, adapters, supervised fine-tuning, and retrieval-augmented generation (RAG). Attendees will learn how to fine-tune LLMs with minimal computational resources while maintaining or enhancing performance, with practical demonstrations using the NVIDIA NeMo Framework. By the end, participants will be equipped to efficiently customize LLMs for specific tasks and challenges. This talk is valuable for researchers, developers, and AI enthusiasts interested in optimizing and personalizing LLMs.

Related talks at this congress

Open session

World Congress 2024

July 18, 2024 · 12:50–13:20

STAGE 9 (600)

Unlocking the Power of AI: Accessible Language Model Tuning for All

Cedric Clyburn, @technicallylegare

Cedric Clyburn
@technicallylegare
Open session

World Congress 2024

July 18, 2024 · 11:30–12:00

STAGE 11 (700)

Efficient deployment and inference of GPU-accelerated LLMs​

Adolf Hohl

Sr. Mgr. Solution Architects AUTO Enterprise

Adolf Hohl
Open session

World Congress 2024

July 19, 2024 · 14:20–14:50

STAGE 6 (120)

Intro to LLMs and recent developments

Christian Winkler

CEO of datanizing, research professor at TH Nürnberg

Christian Winkler
Open session

World Congress 2024

July 18, 2024 · 15:30–16:00

STAGE 3 (120)

Leveraging Large Language Models for Legacy Code Translation: Challenges and Solutions

Michael Niebisch

Software Developer Algorithms ZEISS

Michael Niebisch
All sessions at this congress