2018
Less is More: Picking Informative Frames for Video Captioning
ECCV 2018poster
In video captioning task, the best practice has been achieved by attention-based models which associate salient visual components with sentences in the video. However, existing study follows a common procedure which includes a frame-level appearance modeling and motion modeling on equal interval fra…