← Search

Lorenzo Vaquero

3 accepted papers

2026

From Weights to Concepts: Data-Free Interpretability of CLIP via Singular Vector Decomposition

CVPR 2026

As vision-language models are deployed at scale, understanding their internal mechanisms becomes increasingly critical. Existing interpretability methods predominantly rely on activations, making them dataset-dependent, vulnerable to data bias, and often restricted to coarse head-level explanations.

Cited by 0SourceScholar
2025

ConViS-Bench: Estimating Video Similarity Through Semantic Concepts

NeurIPS 2025poster

What does it mean for two videos to be similar? Videos may appear similar when judged by the actions they depict, yet entirely different if evaluated based on the locations where they were filmed. While humans naturally compare videos by taking different aspects into account, this ability has not be…

Cited by 0SourceScholar
2025

Superpowering Open-Vocabulary Object Detectors for X-ray Vision

ICCV 2025poster

Open-vocabulary object detection (OvOD) is set to revolutionize security screening by enabling systems to recognize any item in X-ray scans. However, developing effective OvOD models for X-ray imaging presents unique challenges due to data scarcity and the modality gap that prevents direct adoption…