2024
GAOKAO-MM: A Chinese Human-Level Benchmark for Multimodal Models Evaluation
ACL 2024findings
The Large Vision-Language Models (LVLMs) have demonstrated great abilities in image perception and language understanding. However, existing datasets either focus solely on primary perception abilities and commonsense knowledge, or have a low level of text comprehension difficulty, which are insuffi…