2023
How to Push the Fastest Model 50x Faster: Streaming Non-Autoregressive Speech Synthesis on Resouce-Limited Devices
ICASSP 2023accepted
Minimizing the latency of the text-to-speech system on end-user resource-limited devices is one of the top demands of voice-based human-machine interaction applications. In this paper, the FastStreamSpeech model is proposed combining the advantages of the advanced approaches in neural-based speech s…