← Search

Zhuotao Tian*

1 accepted papers

2024

Explore the Potential of CLIP for Training-Free Open Vocabulary Semantic Segmentation

ECCV 2024poster

"CLIP, as a vision-language model, has significantly advanced Open-Vocabulary Semantic Segmentation (OVSS) with its zero-shot capabilities. Despite its success, its application to OVSS faces challenges due to its initial image-level alignment training, which affects its performance in tasks requirin…