2023
Confidence-aware Pseudo-label Learning for Weakly Supervised Visual Grounding
ICCV 2023poster
Visual grounding aims at localizing the target object in image which is most related to the given free-form natural language query. As labeling the position of target object is labor-intensive, the weakly supervised methods, where only image-sentence annotations are required during model training ha…