← Search

Weiran Chen

3 accepted papers

2024

TARN-VIST: Topic Aware Reinforcement Network for Visual Storytelling

COLING 2024main

As a cross-modal task, visual storytelling aims to generate a story for an ordered image sequence automatically. Different from the image captioning task, visual storytelling requires not only modeling the relationships between objects in the image but also mining the connections between adjacent im…

Cited by 1SourcePDFScholar