2025
AdaReTaKe: Adaptive Redundancy Reduction to Perceive Longer for Video-language Understanding
ACL 2025finding
Multimodal Large Language Models (MLLMs) have revolutionized video understanding, yet are still limited by context length when processing long videos. Recent methods compress videos by leveraging visual redundancy uniformly, yielding promising results. Nevertheless, our quantitative analysis shows t…