2026
Time Blindness: Why Video-Language Models Can't See What Humans Can?
CVPR 2026
Recent advances in vision-language models (VLMs) have made impressive strides in understanding spatio-temporal relationships in videos. However, when spatial information is obscured, these models struggle to capture purely temporal patterns. We introduce SpookyBench, a benchmark where information is