← Search

Aminreza Sefid

1 accepted papers

2025

CLIP Under the Microscope: A Fine-Grained Analysis of Multi-Object Representation

CVPR 2025poster

Contrastive Language-Image Pre-training (CLIP) models excel in zero-shot classification, yet face challenges in complex multi-object scenarios. This study offers a comprehensive analysis of CLIP's limitations in these contexts using a specialized dataset, ComCO, designed to evaluate CLIP's encoders…