2025
Semi-Supervised CLIP Adaptation by Enforcing Semantic and Trapezoidal Consistency
ICLR 2025poster
Vision-language pre-training models, such as CLIP, have demonstrated strong capability in rapidly adapting to downstream tasks through fine-tuning, and have been widely applied across various tasks. However, when the downstream tasks are constrained by limited image-text paired data, CLIP struggles…