2025
Hierarchical Cross-modal Prompt Learning for Vision-Language Models
ICCV 2025poster
Pre-trained Vision-Language Models (VLMs) such as CLIP have shown excellent generalization abilities. However, adapting these large-scale models to downstream tasks while preserving their generalization capabilities remains challenging. Although prompt learning methods have shown promise, they suffe…