2024
OLKAVS: An Open Large-Scale Korean Audio-Visual Speech Dataset
ICASSP 2024accepted
Inspired by humans comprehending speech in a multi-modal manner, various audio-visual datasets have been constructed. However, most existing datasets focus on English, developed from pre-existing videos using various prediction models, and have only a small number of multi-view videos. To mitigate t…