2024
uCAP: An Unsupervised Prompting Method for Vision-Language Models
ECCV 2024oral
"This paper addresses a significant limitation that prevents Contrastive Language-Image Pretrained Models (CLIP) from achieving optimal performance on downstream image classification tasks. The key problem with CLIP-style zero-shot classification is that it requires domain-specific context in the fo…