World Congress 2024 Aug 20, 2024 Session details

Unlocking the Power of AI: Accessible Language Model Tuning for All

Cedric Clyburn , @technicallylegare

Stop overpaying for massive, generic LLMs. Standard engineering teams can use InstructLab to locally fine-tune open-source models, drastically cutting inference costs without specialized data science expertise.

Pause
Mute Enter Fullscreen
#1 about 4 min

Moving beyond generalist models with local fine tuning

Adapting large language models for specific use cases addresses the limitations of generalist AI.

#2 about 3 min

Overcoming challenges and limitations of generative artificial intelligence

Knowledge cutoffs, lack of transparency, legal exposures, and deployment costs limit enterprise use of foundation models.

#3 about 5 min

Evaluating methods for improving large language model outputs

Prompt engineering, parameter efficient fine tuning, alignment tuning, and retrieval-augmented generation enhance model application performance.

#4 about 4 min

Selecting the appropriate foundation model for enterprise applications

Choosing appropriately sized and permissively licensed foundation models like Granite significantly reduces processing costs and deployment time.

#5 about 4 min

Simplifying open source model fine tuning with InstructLab

A community-driven project enables developers to contribute knowledge and skills to language models using simple YAML structures.

#6 about 9 min

Generating training data and fine tuning with InstructLab

Setting up a local taxonomy framework allows for synthetic data generation and parameter efficient model training.

#7 about 5 min

Integrating trained language models into enterprise Java applications

Serving a locally trained model within a Java web socket application provides tailored responses for specific business domains.

Matching moments

2:37 min

Understanding core parameters and mechanics of large language models

Julián Duque Julián Duque · WWC 2025

2:00 min

Navigating the layers of the language model inference stack

Christin Pohl Christin Pohl · WWC Europe 2026

5:01 min

Leveraging large language models for code optimization and development

Stephan Gillich Stephan Gillich +3 · WWC 2024

4:25 min

Overcoming AI hallucinations and restrictive content guardrails

Perf + AI

1:21 min

Cost considerations of fine-tuning large language models

Kevin Klues Kevin Klues

1:39 min

Balancing human-centric AI collaboration with environmental sustainability practices

Madalena Costa Madalena Costa · WWC 2024

Upcoming sessions on this topic

Open session

World Congress 2026 North America

Fast, Cheap, and Accurate: Optimizing LLM Inference with vLLM and Quantization

Legare Kerrison, Cedric Clyburn

Legare Kerrison
Cedric Clyburn
Open session

World Congress 2026 North America

No Single Model to Rule Them All: Building Resilient AI Agents Across Open & Closed LLMs

Emmanuel Acheampong

Senior Manager Developer Relations at Crusoe AI

Emmanuel Acheampong
Open session

World Congress 2026 North America

Understanding LLM Architectures: Inside the Design of Modern Models

Jofia Jose Prakash

Enterprise AI Architect at American Chemical Society

Jofia Jose Prakash
Open session

World Congress 2026 North America

You Can’t Re-Run Sunlight: Designing ML Data Architectures for Physical AI

An Phan

Senior Data Infrastructure Engineer @ Hippo Harvest

An Phan
Open session

World Congress 2026 North America

Building Stuff with GenAI - The Open Minded Workshop beyond OpenAI

Andreas Erben

CTO for Applied AI and Metaverse at daenet

Andreas Erben
Open session

World Congress 2026 North America

Making Science Larger, not just Faster

Yuval Dvir

Commercial Executive, SandboxAQ

Yuval Dvir