← Search

Yuanjing Luo

1 accepted papers

2025

ALLVB: All-in-One Long Video Understanding Benchmark

AAAI 2025technical

From image to video understanding, the capabilities of Multi-modal LLMs (MLLMs) are increasingly powerful. However, most existing video understanding benchmarks are relatively short, which makes them inadequate for effectively evaluating the long-sequence modeling capabilities of MLLMs. This highlig…

Cited by 0SourcePDFScholar