2026
LearnPruner: Rethinking Attention-based Token Pruning in Vision Language Models
ICLR 2026poster
Vision-Language Models (VLMs) have recently demonstrated remarkable capabilities in visual understanding and reasoning, but they also impose significant computational burdens due to long visual sequence inputs. Recent works address this issue by pruning unimportant visual tokens, achieving substanti…