2024
Unified Embedding Alignment for Open-Vocabulary Video Instance Segmentation
ECCV 2024poster
"Open-Vocabulary Video Instance Segmentation (VIS) is attracting increasing attention due to its ability to segment and track arbitrary objects. However, the recent Open-Vocabulary VIS attempts obtained unsatisfactory results, especially in terms of generalization ability of novel categories. We dis…