2026
HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models
ICLR 2026poster
Multimodal Large Language Models (MLLMs) have demonstrated significant potential to advance a broad range of domains. However, current benchmarks for evaluating MLLMs primarily emphasize general knowledge and vertical step-by-step reasoning typical of STEM disciplines, while overlooking the distinct…