← Search

Zhongtian Fu

1 accepted papers

2024

Noise-Aware Image Captioning with Progressively Exploring Mismatched Words

AAAI 2024technical

Image captioning aims to automatically generate captions for images by learning a cross-modal generator from vision to language. The large amount of image-text pairs required for training is usually sourced from the internet due to the manual cost, which brings the noise with mismatched relevance th…