← Search

Georgii Mikriukov

2 accepted papers

2026

Explaining, Verifying and Aligning Semantic Hierarchies in Vision-Language Model Embeddings

IJCAI 2026

Vision-language model (VLM) encoders such as CLIP enable strong retrieval and zero-shot classification in a shared image–text embedding space, yet the semantic organization of this space is rarely inspected. We present a post-hoc framework to explain, verify, and align the semantic hierarchies induc

Cited by 0Scholar
2022

Unsupervised Contrastive Hashing for Cross-Modal Retrieval in Remote Sensing

ICASSP 2022accepted

The development of cross-modal retrieval systems that can search and retrieve semantically relevant data across different modalities based on a query in any modality has attracted great attention in remote sensing (RS). In this paper, we focus our attention on cross-modal text-image retrieval, where…

Cited by 0SourceScholar