2023
VL-Match: Enhancing Vision-Language Pretraining with Token-Level and Instance-Level Matching
ICCV 2023poster
Vision-Language Pretraining (VLP) has significantly improved the performance of various vision-language tasks with the matching of images and texts. In this paper, we propose VL-Match, a Vision-Language framework with Enhanced Token-level and Instance-level Matching. At the token level, a Vision-Lan…