← Search

Li Cao

1 accepted papers

2025

AdaReTaKe: Adaptive Redundancy Reduction to Perceive Longer for Video-language Understanding

ACL 2025finding

Multimodal Large Language Models (MLLMs) have revolutionized video understanding, yet are still limited by context length when processing long videos. Recent methods compress videos by leveraging visual redundancy uniformly, yielding promising results. Nevertheless, our quantitative analysis shows t…