2023
Divcon: Learning Concept Sequences for Semantically Diverse Image Captioning
ICASSP 2023accepted
Human generated image captions contain diverse semantic concepts, while this is still a difficult task for machines. The frequency distribution of semantic concepts in datasets is usually extremely imbalanced, leading to models repeatedly describe frequently occurring semantic concepts, resulting in…