2025
InfiniBench: A Benchmark for Large Multi-Modal Models in Long-Form Movies and TV Shows
EMNLP 2025
Understanding long-form videos, such as movies and TV episodes ranging from tens of minutes to two hours, remains a significant challenge for multi-modal models. Existing benchmarks often fail to test the full range of cognitive skills needed to process these temporally rich and narratively complex