2025
Don’t Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models
ACL 2025finding
Large Vision Language Models (LVLMs) demonstrate strong capabilities in visual understanding and description, yet often suffer from hallucinations, attributing incorrect or misleading features to images. We observe that LVLMs disproportionately focus on a small subset of image tokens—termed blind to…