2026
TOP-RL: Task-Optimized Progressive Token Pruning with Reinforcement Learning for Vision Language Models
AAAI 2026technical
In recent years, Large Vision-Language Models (LVLMs) have significantly advanced multimodal tasks. However, their inference requires intensive processing of numerous visual tokens and incurs substantial computational overhead. Existing methods typically compress visual tokens either at the input st