2025
AdaV: Adaptive Text-visual Redirection for Vision-Language Models
ACL 2025finding
The success of Vision-Language Models (VLMs) often relies on high-resolution schemes that preserve image details, while these approaches also generate an excess of visual tokens, leading to a substantial decrease in model efficiency. A typical VLM includes a visual encoder, a text encoder, and an LL…