2024
SegVG: Transferring Object Bounding Box to Segmentation for Visual Grounding
ECCV 2024poster
"Different from Object Detection, Visual Grounding deals with detecting a bounding box for each text-image pair. This one box for each text-image data provides sparse supervision signals. Although previous works achieve impressive results, their passive utilization of annotation, i.e. the sole use o…