2024
MMBENCH: Is Your Multi-Modal Model an All-around Player?
ECCV 2024oral
"Large vision-language models (VLMs) have recently achieved remarkable progress, exhibiting impressive multimodal perception and reasoning abilities. However, effectively evaluating these large VLMs remains a major challenge, hindering future development in this domain. Traditional benchmarks like V…