2025
Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations
ICLR 2025poster
We investigate the internal representations of vision-language models (VLMs) to address hallucinations, a persistent challenge despite advances in model size and training. We project VLMs’ internal image representations to their language vocabulary and observe more confident output probabilities on…