← Search

Wuliang Huang

4 accepted papers

2025

Aligning Composed Query with Image via Discriminative Perception from Negative Correspondences

AAAI 2025technical

The task of composed image retrieval aims to match the multi-modal query composed of a reference image and a modification sentence with the target image. Most current approaches narrow the distances between the composed queries and targets by investigating matched correspondences in positive triplet…

Cited by 0SourcePDFScholar
2025

Mitigating Pervasive Modality Absence Through Multimodal Generalization and Refinement

AAAI 2025technical

The performance of multimodal models often deteriorates when modality absence occurs. The absence disrupts the learned inter-modal correlations, resulting in biased multimodal representations. This challenge is especially pronounced when the absence is pervasive, affecting both the training and infe…

Cited by 0SourcePDFScholar
2025

Towards Robust Uncertainty Calibration for Composed Image Retrieval

NeurIPS 2025poster

The interactive task of composed image retrieval aims to retrieve the most relevant images with the bi-modal query, consisting of a reference image and a modification sentence. Despite significant efforts to bridge the heterogeneous gap within the bi-modal query and leverage contrastive learning to…

Cited by 0SourceScholar
2025

VersaFusion: A Versatile Diffusion-Based Framework for Fine-Grained Image Editing and Enhancement

AAAI 2025technical

Text-to-image (T2I) diffusion models have achieved remarkable progress in generating realistic images from textual descriptions. However, ensuring consistent high-quality image generation with complete backgrounds, object appearance, and optimal texture rendering remains challenging. This paper pres…