2023
Learning Cross-Modal Affinity for Referring Video Object Segmentation Targeting Limited Samples
ICCV 2023poster
Referring video object segmentation (RVOS), as a supervised learning task, relies on sufficient annotated data for a given scene. However, in more realistic scenarios, only minimal annotations are available for a new scene, which poses significant challenges to existing RVOS methods. With this in mi…