2025
Where is this coming from? Making groundedness count in the evaluation of Document VQA models
NAACL 2025findings
Document Visual Question Answering (VQA) models have evolved at an impressive rate over the past few years, coming close to or matching human performance on some benchmarks. We argue that common evaluation metrics used by popular benchmarks do not account for the semantic and multimodal groundedness…