2025
DocAgent: An Agentic Framework for Multi-Modal Long-Context Document Understanding
EMNLP 2025
Recent advances in large language models (LLMs) have demonstrated significant promise in document understanding and question-answering. Despite the progress, existing approaches can only process short documents due to limited context length or fail to fully leverage multi-modal information. In this