2024
Investigating the Clusters Discovered By Pre-Trained AV-HuBERT
ICASSP 2024accepted
Self-supervised models, such as HuBERT and its audio-visual version AV-HuBERT, have demonstrated excellent performance on various tasks. The main factor for their success is the pre-training procedure, which requires only raw data without human transcription. During the self-supervised pre-training…