IJCAI 2021poster26 citations

Norm-guided Adaptive Visual Embedding for Zero-Shot Sketch-Based Image Retrieval

Wenjie Wang, Yufeng Shi, Shiming Chen, Qinmu Peng, Feng Zheng, Xinge You

Abstract

Zero-shot sketch-based image retrieval (ZS-SBIR), which aims to retrieve photos with sketches under the zero-shot scenario, has shown extraordinary talents in real-world applications. Most existing methods leverage language models to generate class-prototypes and use them to arrange the locations of all categories in the common space for photos and sketches. Although great progress has been made, few of them consider whether such pre-defined prototypes are necessary for ZS-SBIR, where locations of unseen class samples in the embedding space are actually determined by visual appearance and a visual embedding actually performs better. To this end, we propose a novel Norm-guided Adaptive Visual Embedding (NAVE) model, for adaptively building the common space based on visual similarity instead of language-based pre-defined prototypes. To further enhance the representation quality of unseen classes for both photo and sketch modality, modality norm discrepancy and noisy label regularizer are jointly employed to measure and repair the modality bias of the learned common embedding. Experiments on two challenging datasets demonstrate the superiority of our NAVE over state-of-the-art competitors.

Computer Vision: Recognition: Detection, Categorization, Indexing, Matching, Retrieval, Semantic InterpretationMachine Learning: Deep LearningMachine Learning: Multi-instanceMulti-labelMulti-view learning
BibTeX
@inproceedings{ijcai2021p153,
  title     = {Norm-guided Adaptive Visual Embedding for Zero-Shot Sketch-Based Image Retrieval},
  author    = {Wang, Wenjie and Shi, Yufeng and Chen, Shiming and Peng, Qinmu and Zheng, Feng and You, Xinge},
  booktitle = {Proceedings of the Thirtieth International Joint Conference on
               Artificial Intelligence, {IJCAI-21}},
  publisher = {International Joint Conferences on Artificial Intelligence Organization},
  editor    = {Zhi-Hua Zhou},
  pages     = {1106--1112},
  year      = {2021},
  month     = {8},
  note      = {Main Track},
  doi       = {10.24963/ijcai.2021/153},
  url       = {https://doi.org/10.24963/ijcai.2021/153},
}
Norm-guided Adaptive Visual Embedding for Zero-Shot Sketch-Based Image Retrieval · IJCAI 2021