2025
Semantic and Sequential Alignment for Referring Video Object Segmentation
CVPR 2025poster
Referring video object segmentation (RVOS) seeks to segment the objects within a video referred by linguistic expressions. Existing RVOS solutions follow a "fuse then select" paradigm: establishing semantic correlation between visual and linguistic feature, and performing frame-level query interacti…