2024
DEAL: Disentangle and Localize Concept-level Explanations for VLMs
ECCV 2024poster
"Large pre-trained Vision-Language Models (VLMs) have become ubiquitous foundational components of other models and downstream tasks. Although powerful, our empirical results reveal that such models might not be able to identify fine-grained concepts. Specifically, the explanations of VLMs with resp…