2026
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs
ICML 2026poster
Recent advances in Omni-Multimodal Large Language Models (Omni-MLLMs) have enabled strong integration of vision, audio, and language. However, their audio-visual intelligence (AVI) remains insufficiently evaluated due to the lack of systematic and comprehensive benchmarks. We introduce AVI-Bench, a …