← Search

Qiyuan Fan

1 accepted papers

2025

DocAssistant: Integrating Key-region Reading and Step-wise Reasoning for Robust Document Visual Question Answering

EMNLP 2025

Understanding the multimodal documents is essential for accurately extracting relevant evidence and using it for reasoning. Existing document understanding models struggle to focus on key information and tend to generate answers straightforwardly, ignoring evidence from source documents and lacking

Cited by 0SourcePDFScholar