2025
Advancing Interpretability of CLIP Representations with Concept Surrogate Model
NeurIPS 2025poster
Contrastive Language-Image Pre-training (CLIP) generates versatile multimodal embeddings for diverse applications, yet the specific information captured within these representations is not fully understood. Current explainability techniques often target specific tasks, overlooking the rich, general…