2026
MosaicDoc: A Large-Scale Bilingual Benchmark for Visually Rich Document Understanding
AAAI 2026technical
Despite the rapid progress of Vision-Language Models (VLMs), their capabilities are inadequately assessed by existing benchmarks, which are predominantly English-centric, feature simplistic layouts, and support limited tasks. Consequently, they fail to evaluate model performance for Visually Rich Do