← Search

Sangyeon Cho

1 accepted papers

2025

Learning to See through Sound: From VggCaps to Multi2Cap for Richer Automated Audio Captioning

EMNLP 2025

Automated Audio Captioning (AAC) aims to generate natural language descriptions of audio content, enabling machines to interpret and communicate complex acoustic scenes. However, current AAC datasets often suffer from short and simplistic captions, limiting model expressiveness and semantic depth. T

Cited by 0SourcePDFScholar