2026
StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding
ICASSP 2026poster
The rapid development of Multimodal Large Language Models (MLLMs) has expanded their capabilities from image comprehension to video understanding. However, most of these MLLMs focus primarily on offline video comprehension, necessitating extensive processing of all video frames before any queries ca…