2022
Self Supervised Representation Learning with Deep Clustering for Acoustic Unit Discovery from Raw Speech
ICASSP 2022accepted
The automatic discovery of acoustic sub-word units from raw speech, without any text or labels, is a growing field of research. The key challenge is to derive representations of speech that can be categorized into a small number of phoneme-like units which are speaker invariant and can broadly captu…