← Search

Delong Zeng

2 accepted papers

2025

Enhancing Multimodal Retrieval via Complementary Information Extraction and Alignment

ACL 2025long

Multimodal retrieval has emerged as a promising yet challenging research direction in recent years. Most existing studies in multimodal retrieval focus on capturing information in multimodal data that is similar to their paired texts, but often ignores the complementary information contained in mult…

2025

Zero-Shot Image Captioning with Multi-type Entity Representations

AAAI 2025technical

As data and computational resources continue to expand, incorporating a variety of knowledge during the pre-training phase enhances large models, providing them with strong zero-shot capabilities. Due to the alignment of modal features by visual language models, zero-shot image captioning no longer…

Cited by 0SourcePDFScholar