2024
Empowering Vision-Language Models for Reasoning Ability through Large Language Models
ICASSP 2024accepted
Vision-language models (VLM) have shown excellent performance in vision-language tasks. However, they sometimes lack sufficient reasoning ability. In contrast, large language models (LLMs) have emerged with powerful reasoning capabilities. Therefore, we propose a framework called TReE, which transfe…