Adolf Hohl
Builds efficient AI infrastructure and inference solutions for automotive enterprises.
Adolf designs AI infrastructure for automotive enterprise customers across Europe. He works to ensure AI workloads run efficiently at scale, building out the systems required to train and deploy models in modern datacenters.
Much of his work centers on inferencing and GPU-accelerated large language models. During his talk on efficient LLM deployment, he broke down how to serve models across multi-GPU environments using TensorRT-LLM and Triton Inference Server. He likes taking inference solutions out of the lab and seeing them operate in the real world.