2024
Open-Set Recognition in the Age of Vision-Language Models
ECCV 2024poster
"Are vision-language models (VLMs) for open-vocabulary perception inherently open-set models because they are trained on internet-scale datasets? We answer this question with a clear no – VLMs introduce closed-set assumptions via their finite query set, making them vulnerable to open-set conditions.…