2024
ArtVLM: Attribute Recognition Through Vision-Based Prefix Language Modeling
ECCV 2024poster
"Recognizing and disentangling visual attributes from objects is a foundation to many computer vision applications. While large vision-language representations like CLIP had largely resolved the task of zero-shot object recognition, zero-shot visual attribute recognition remains a challenge because…