2026
MME-SCI: A Comprehensive and Challenging Science Benchmark for Multimodal Large Language Models
AAAI 2026technical
Recently, multimodal large language models (MLLMs) have achieved significant advancements across various domains, and corresponding evaluation benchmarks have been continuously refined and improved. In this process, benchmarks in the scientific domain have played an important role in assessing the r