← Search

Gesina Schwalbe

2 accepted papers

2026

Explaining, Verifying and Aligning Semantic Hierarchies in Vision-Language Model Embeddings

IJCAI 2026

Vision-language model (VLM) encoders such as CLIP enable strong retrieval and zero-shot classification in a shared image–text embedding space, yet the semantic organization of this space is rarely inspected. We present a post-hoc framework to explain, verify, and align the semantic hierarchies induc

Cited by 0Scholar