ICASSP 2022accepted0 citations

Semantic Association Network for Video Corpus Moment Retrieval

Dahyun Kim, Sunjae Yoon, Ji Woo Hong, Chang D. Yoo

Abstract

This paper considers Semantic Association Network (SAN) for Video Corpus Moment Retrieval (VCMR) which localizes temporal moment that best corresponds to the given text query in a corpus of videos. Collaborations among common semantics from multi-modal inputs are essential for effectively understanding video together with subtitle and text query. For this collaboration, SAN associates common semantics within the same modality (by Intra Semantic Association) and across different modalities (by Inter Semantic Association) with dedicated module referred to as Modality Semantic Association (MSA). SAN surpasses existing state-of-the-art performance on the TVR and DiDeMo benchmark datasets. Extensive ablation studies and qualitative analyses show the effectiveness of the proposed model.

BibTeX
@inproceedings{icassp2022_semanticassociat,
  title = {Semantic Association Network for Video Corpus Moment Retrieval},
  author = {Dahyun Kim and Sunjae Yoon and Ji Woo Hong and Chang D. Yoo},
  booktitle = {ICASSP 2022},
  year = {2022}
}