← Search

Liyuan Jiang

1 accepted papers

2026

Token Reduction via Local and Global Contexts Optimization for Efficient Video Large Language Models

CVPR 2026

Video Large Language Models (VLLMs) demonstrate strong video understanding but suffer from inefficiency due to redundant visual tokens. Existing pruning primary targets intra-frame spatial redundancy or prunes inside the LLM with shallow-layer overhead, yielding suboptimal spatiotemporal reduction a

Cited by 0SourcecodeScholar