2022
Informative Attention Supervision for Grounded Video Description
ICASSP 2022accepted
Attention supervision encourages grounded video description models (GVDMs) to focus on the related visual content when generating words. Thus, it improves the description performance of GVDMs. However, existing GVDMs often fail to focus on small but informative regions because these regions are cons…