2025
Incorporating Dense Knowledge Alignment into Unified Multimodal Representation Models
CVPR 2025poster
Leveraging Large Language Models (LLMs) for text representation has achieved significant success, but the exploration of using Multimodal LLMs (MLLMs) for multimodal representation remains limited. Previous MLLM-based representation studies have primarily focused on unifying the embedding space whil…