World Congress 2025
July 11, 2025 · 13:00–13:30
Stage 4
Prompt Injection, Poisoning & More: The Dark Side of LLMs
Keno Dreßel
Principal Consultant & Head of AI @ SQUER
World Congress 2025
Recent studies in 2024 have revolutionised our understanding of large language models (LLMs).
This talk explores three key discoveries.
First, research shows Llama 2 models use English as their internal representation regardless of input/output language, explaining certain biases.
Second, breakthroughs by Anthropic and OpenAI have revealed monosemantic features in Claude 3 and GPT-4, enabling better understanding and adjustment of topic-specific behaviours.
Third, studies demonstrate why LLMs memorise outlier data, particularly unique strings and personal information, explaining instances of privacy breaches. We’ll discuss implications for LLM privacy and security.
Attendees will walk away with a deeper understanding of the inner workings of LLMs, and with hints to mitigate their intrinsic limitations.
World Congress 2025
July 11, 2025 · 13:00–13:30
Stage 4
Keno Dreßel
Principal Consultant & Head of AI @ SQUER
World Congress 2025
July 11, 2025 · 13:55–14:05
Airstream 1
Riccardo Zoncada
Software Engineer @ xtream
World Congress 2025
July 10, 2025 · 10:50–11:20
Stage 11
Tomislav Tipurić
Chief Technology Officer, Nephos
World Congress 2025
July 10, 2025 · 12:50–13:20
Stage 2
Simon A.T. Jiménez
create better software specifications in less time