← Search

Grégoire Dhimoïla

2 accepted papers

2026

Cross-Modal Redundancy and the Geometry of Vision–Language Embeddings

ICLR 2026poster

Vision–language models (VLMs) align images and text with remarkable success, yet the geometry of their shared embedding space remains poorly understood. To probe this geometry, we begin from the Iso-Energy Assumption, which exploits cross-modal redundancy: a concept that is truly shared should exhi…

Cited by 0SourceScholar
2024

Cluster-Norm for Unsupervised Probing of Knowledge

EMNLP 2024main

The deployment of language models brings challenges in generating reliable text, especially when these models are fine-tuned with human preferences. To extract the encoded knowledge in these models without (potentially) biased human labels, unsupervised probing techniques like Contrast-Consistent Se…