2024
Multi-Label Cluster Discrimination for Visual Representation Learning
ECCV 2024poster
"Contrastive Language Image Pre-training (CLIP) has recently demonstrated success across various tasks due to superior feature representation empowered by image-text contrastive learning. However, the instance discrimination method used by CLIP can hardly encode the semantic structure of training da…