← Search

Shaojun E

1 accepted papers

2025

MAGE: Multimodal Alignment and Generation Enhancement via Bridging Visual and Semantic Spaces

IJCAI 2025

In the latest advancements in multimodal learning, effectively addressing the spatial and semantic losses of visual data after encoding remains a critical challenge. This is because the performance of large multimodal models is positively correlated with the coupling between visual encoders and larg