← Search

Haoxuan Cheng

1 accepted papers

2026

StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding

ICASSP 2026poster

The rapid development of Multimodal Large Language Models (MLLMs) has expanded their capabilities from image comprehension to video understanding. However, most of these MLLMs focus primarily on offline video comprehension, necessitating extensive processing of all video frames before any queries ca…

Cited by 0SourcePDFScholar