2025
VLMs have Tunnel Vision: Evaluating Nonlocal Visual Reasoning in Leading VLMs
NeurIPS 2025spotlight
Vision Language Models (VLMs) excel at complex visual tasks such as VQA and chart understanding, yet recent work suggests they struggle with simple perceptual tests. We present an evaluation that tests vision-language models’ capacity for \emph{nonlocal visual reasoning}- reasoning that requires cha…