← Search

Jiasun Feng

1 accepted papers

2026

Robust Cross-Modal Retrieval via Generative Semantic Refinement and Exclusion-Guided Adaptation

ICML 2026poster

Vision-Language Pre-trained (VLP) models are vulnerable to real-world query noise. Current cross-modal Test-Time Adaptation (TTA) methods often rely on high-confidence predictions, which induces confirmation bias and neglects the informative signals in ambiguous Low-Confidence Queries. To address th…

Cited by 0SourceScholar