← Search

Lorenzo Agnolucci

3 accepted papers

2025

Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion

ICLR 2025poster

Pre-trained multi-modal Vision-Language Models like CLIP are widely used off-the-shelf for a variety of applications. In this paper, we show that the common practice of individually exploiting the text or image encoders of these powerful multi-modal models is highly suboptimal for intra-modal tasks…

2025

Image Intrinsic Scale Assessment: Bridging the Gap Between Quality and Resolution

ICCV 2025poster

Image Quality Assessment (IQA) measures and predicts perceived image quality by human observers. Although recent studies have highlighted the critical influence that variations in the scale of an image have on its perceived quality, this relationship has not been systematically quantified.To bridge…

2023

Zero-Shot Composed Image Retrieval with Textual Inversion

ICCV 2023poster

Composed Image Retrieval (CIR) aims to retrieve a target image based on a query composed of a reference image and a relative caption that describes the difference between the two images. The high effort and cost required for labeling datasets for CIR hamper the widespread usage of existing methods,…

Cited by 126PDFcodeScholar