2025
MM-R3: On (In-)Consistency of Vision-Language Models (VLMs)
ACL 2025finding
With the advent of LLMs and variants, a flurry of research has emerged, analyzing the performance of such models across an array of tasks. While most studies focus on evaluating the capabilities of state-of-the-art (SoTA) Vision Language Models (VLMs) through task accuracy (e.g., visual question ans…