2025
Learning Rich Speech Representations with Acoustic-Semantic Factorization
ICASSP 2025accepted
Self-supervised pretraining has transformed speech representation learning, enabling models to generalize across various downstream tasks. However, empirical studies have highlighted two notable gaps. First, different speech tasks require varying levels of acoustic and semantic information, which ar…