2026
CoIn: Coverage and Informativeness-Guided Token Reduction for Efficient Large Multimodal Models
CVPR 2026
Large Multimodal Models (LMMs) have shown remarkable success in visual understanding tasks. LMMs encode visual and textual inputs into tokens, which are then processed by Large Language Models (LLMs). However, the large number of visual tokens poses a major bottleneck for inference efficiency and me