← Search

Qian-Wei Wang

4 accepted papers

2026

Imagine with Layout and Sketch: Enhancing Vision-Language Retrieval with Dual-Stream Multi-Modal Query Refinement

AAAI 2026technical

Vision-Language Retrieval (VLR) aims to retrieve relevant visual or textual information from multimodal data using language or image queries. However, traditional VLR methods often rely on data-driven shallow semantic alignment and fail to understand the deeper structural and fine-grained entity fea

Cited by 0SourcePDFScholar
2025

Pre-Trained Vision-Language Models as Noisy Partial Annotators

AAAI 2025technical

In noisy partial label learning, each training sample is associated with a set of candidate labels, and the ground-truth label may be contained within this set. With the emergence of powerful pre-trained vision-language models, e.g. CLIP, it is natural to consider using these models to automatically…

2024

Controller-Guided Partial Label Consistency Regularization with Unlabeled Data

AAAI 2024technical

Partial label learning (PLL) learns from training examples each associated with multiple candidate labels, among which only one is valid. In recent years, benefiting from the strong capability of dealing with ambiguous supervision and the impetus of modern data augmentation methods, consistency regu…

Cited by 3SourcePDFScholar
2023

Combating Unknown Bias with Effective Bias-Conflicting Scoring and Gradient Alignment

AAAI 2023technical

Models notoriously suffer from dataset biases which are detrimental to robustness and generalization. The identify-emphasize paradigm shows a promising effect in dealing with unknown biases. However, we find that it is still plagued by two challenges: A, the quality of the identified bias-conflictin…

Cited by 9SourcePDFScholar