2025
LLaVA-KD: A Framework of Distilling Multimodal Large Language Models
ICCV 2025poster
The success of Large Language Models (LLMs) has inspired the development of Multimodal Large Language Models (MLLMs) for unified understanding of vision and language. However, the increasing model size and computational complexity of large-scale MLLMs (l-MLLMs) limit their use in resource-constraine…