← Search

Yu-Jen Tseng

1 accepted papers

2025

Bridging Compressed Image Latents and Multimodal Large Language Models

ICLR 2025poster

This paper presents the first-ever study of adapting compressed image latents to suit the needs of downstream vision tasks that adopt Multimodal Large Language Models (MLLMs). MLLMs have extended the success of large language models to modalities (e.g. images) beyond text, but their billion scale hi…

Cited by 1SourcePDFScholar