← Search

Bailing Zhang

3 accepted papers

2023

Fine-Grained Image-Text Matching by Cross-Modal Hard Aligning Network

CVPR 2023poster

Current state-of-the-art image-text matching methods implicitly align the visual-semantic fragments, like regions in images and words in sentences, and adopt cross-attention mechanism to discover fine-grained cross-modal semantic correspondence. However, the cross-attention mechanism may bring redun…

2020

Attentive Prototype Few-shot Learning with Capsule Network-based Embedding

ECCV 2020poster

Few-shot learning, namely recognizing novel categories with a very small amount of training examples, is a challenging area of machine learning research. Traditional deep learning methods require massive training data to tune the huge number of parameters, which is often impractical and prone to ove…

Cited by 51SourcePDFScholar