← Search

Linxin Liu

1 accepted papers

2024

Comprehensive Visual Grounding for Video Description

AAAI 2024technical

The grounding accuracy of existing video captioners is still behind the expectation. The majority of existing methods perform grounded video captioning on sparse entity annotations, whereas the captioning accuracy often suffers from degenerated object appearances on the annotated area such as motion…

Cited by 2SourcePDFScholar