2024
Explore the Potential of CLIP for Training-Free Open Vocabulary Semantic Segmentation
ECCV 2024poster
"CLIP, as a vision-language model, has significantly advanced Open-Vocabulary Semantic Segmentation (OVSS) with its zero-shot capabilities. Despite its success, its application to OVSS faces challenges due to its initial image-level alignment training, which affects its performance in tasks requirin…