2025
PACT: Pruning and Clustering-Based Token Reduction for Faster Visual Language Models
CVPR 2025poster
Visual Language Models require substantial computational resources for inference due to the additional input tokens needed to represent visual information. However, these visual tokens often contain redundant and unimportant information, resulting in an unnecessarily high number of tokens. To addres…