Guest Lecture

How to Encode World Knowledge?

Guest Lecture by Ass. Prof. Dr. Guy Gilboa from EE Dept. (Technion), IIT Haifa
September 24th, 2026
10:00
IMBIT, NEXUS Lab, George Köhler Allee 201, 79110 Freiburg
Foundation models are a key platform which implicitly encodes world knowledge. In this talk we first focus on vision-language models, such as CLIP, and investigate their geometric behavior and logic behind the high-dimensional feature encoding. For instance, we find that as image or text become more rare and distinct they are encoded further from the center of the embedding. We explain why InfoNCE loss leads to that behavior. We also find out empirically that each modality can be well modeled statistically as admitting a multivariate Gaussian distribution. This finding is then proved formally, where it is shown that InfoNCE asymptotically induces Gaussian distribution. Finally, we connect two major image foundation models – encoders and generators, through a universal normal embedding hypothesis. A surprising consequence of this hypothesis is demonstrated, where generative diffusion “noise” contains semantic data which can be accessed by linear probing for either classification or editing. This talk is based on 6 recent papers in the group published at ICML 2025, ICLR 2026 and CVPR 2026.

Prof. Guy Gilboa is a faculty member at the Electrical and Computer Engineering Department, Technion - Israel Institute of Technology since 2013. Previously he has been a postdoctoral fellow at UCLA and a senior researcher at Microsoft and Philips Healthcare. His main research areas are mathematical image processing and machine learning, focusing on representation learning of foundation models.

Host: Prof. Thomass Brox, Computer Vision Group

We hope to see you there!