← Search

Zhaoheng Zheng

4 accepted papers

2024

Large Language Models are Good Prompt Learners for Low-Shot Image Classification

CVPR 2024poster

Low-shot image classification where training images are limited or inaccessible has benefited from recent progress on pre-trained vision-language (VL) models with strong generalizability e.g. CLIP. Prompt learning methods built with VL models generate text features from the class names that only hav…

2024

SEAS: ShapE-Aligned Supervision for Person Re-Identification

CVPR 2024poster

We introduce SEAS using ShapE-Aligned Supervision to enhance appearance-based person re-identification. When recognizing an individual's identity existing methods primarily rely on appearance which can be influenced by the background environment due to a lack of body shape awareness. Although some m…

Cited by 8SourcePDFScholar
2022

FashionVLP: Vision Language Transformer for Fashion Retrieval With Feedback

CVPR 2022poster

Fashion image retrieval based on a query pair of reference image and natural language feedback is a challenging task that requires models to assess fashion related information from visual and textual modalities simultaneously. We propose a new vision-language transformer based model, FashionVLP, tha…

Cited by 120PDFScholar
2022

Self-Supervised Learning for Sentiment Analysis via Image-Text Matching

ICASSP 2022accepted

There is often a resemblance in the sentiment expressed in social media posts (text) and their accompanying images. In this paper, We leverage this sentiment congruence for self-supervised representation learning for sentiment analysis. By teaching the model to pair an image with its corresponding s…

Cited by 0SourceScholar