2024
Data-Efficient Contrastive Language-Image Pretraining: Prioritizing Data Quality over Quantity
AISTATS 2024poster
Contrastive Language-Image Pre-training (CLIP) on large-scale image-caption datasets learns representations that can achieve remarkable zero-shot generalization. However, such models require a massive amount of pre-training data. Improving the quality of the pre-training data has been shown to be mu…