2026
HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models
ICML 2026poster
Large vision-language models can produce object hallucinations in image descriptions, highlighting the need for effective detection and mitigation strategies. Prior work commonly relies on the model's attention weights on visual tokens as a detection signal. We reveal that coarse-grained attention-b…