2023
Leveraging Weighted Cross-Graph Attention for Visual and Semantic Enhanced Video Captioning Network
AAAI 2023technical
Video captioning has become a broad and interesting research area. Attention-based encoder-decoder methods are extensively used for caption generation. However, these methods mostly utilize the visual attentive feature to highlight the video regions while overlooked the semantic features of the avai…