2026
Multimodal Aligned Semantic Knowledge for Unpaired Image-text Matching
ICLR 2026oral
While existing approaches address unpaired image-text matching by constructing cross-modal aligned knowledge, they often fail to identify semantically corresponding visual representations for Out-of-Distribution (OOD) words. Moreover, the distributional variance of visual representations associated…