2025
Video-ColBERT: Contextualized Late Interaction for Text-to-Video Retrieval
CVPR 2025poster
In this work, we tackle the problem of text-to-video retrieval (T2VR). Inspired by the success of late interaction techniques in text-document, text-image, and text-video retrieval, our approach, Video-ColBERT, introduces a simple and efficient mechanism for fine-grained similarity assessment betwee…