← Search

Yang-Yang Liu

1 accepted papers

2025

Recoverable Compression: A Multimodal Vision Token Recovery Mechanism Guided by Text Information

AAAI 2025technical

With the advancement of large-scale language modeling techniques, large multimodal models combining visual encoders with large language models have demonstrated exceptional performance in various visual tasks. Most of the current large multimodal models achieve this by mapping visual features obtain…