2025
METok: Multi-Stage Event-based Token Compression for Efficient Long Video Understanding
EMNLP 2025
Recent advances in Video Large Language Models (VLLMs) have significantly enhanced their ability to understand video content. Nonetheless, processing long videos remains challenging due to high computational demands and the redundancy present in the visual data. In this work, we propose METok , a tr