2024
In Defense of Lazy Visual Grounding for Open-Vocabulary Semantic Segmentation
ECCV 2024poster
"We present Lazy Visual Grounding for open-vocabulary semantic segmentation, which decouples unsupervised object mask discovery from object grounding. Plenty of the previous art casts this task as pixel-to-text classification without object-level comprehension, leveraging the image-to-text classific…