2022
On the Interplay between Sparsity, Naturalness, Intelligibility, and Prosody in Speech Synthesis
ICASSP 2022accepted
Are end-to-end text-to-speech (TTS) models over-parametrized? To what extent can these models be pruned, and what happens to their synthesis capabilities? This work serves as a starting point to explore pruning both spectrogram prediction networks and vocoders. We thoroughly investigate the tradeoff…