2023
Unsupervised Pre-Training for Data-Efficient Text-to-Speech on Low Resource Languages
ICASSP 2023accepted
Neural text-to-speech (TTS) models can synthesize natural human speech when trained on large amounts of transcribed speech. How-ever, collecting such large-scale transcribed data is expensive. This paper proposes an unsupervised pre-training method for a sequence-to-sequence TTS model by leveraging…