2025
Mixed Signals: Decoding VLMs’ Reasoning and Underlying Bias in Vision-Language Conflict
EMNLP 2025
Vision-language models (VLMs) have demonstrated impressive performance by effectively integrating visual and textual information to solve complex tasks. However, it is not clear how these models reason over the visual and textual data together, nor how the flow of information between modalities is s