2025
MTabVQA: Evaluating Multi-Tabular Reasoning of Language Models in Visual Space
EMNLP 2025
Vision-Language Models (VLMs) have demonstrated remarkable capabilities in interpreting visual layouts and text. However, a significant challenge remains in their ability to interpret robustly and reason over multi-tabular data presented as images, a common occurrence in real-world scenarios like we