2023
Evaluating Speech-Phoneme Alignment and its Impact on Neural Text-To-Speech Synthesis
ICASSP 2023accepted
In recent years, the quality of text-to-speech (TTS) synthesis vastly improved due to deep-learning techniques, with parallel architectures, in particular, providing excellent synthesis quality at fast inference. Training these models usually requires speech recordings, corresponding phoneme-level t…