2025
Latent Expression Generation for Referring Image Segmentation and Grounding
ICCV 2025poster
Visual grounding tasks, such as referring image segmentation (RIS) and referring expression comprehension (REC), aim to localize a target object based on a given textual description. The target object in an image can be described in multiple ways, reflecting diverse attributes such as color, positio…