2018
Emphatic Speech Prosody Prediction with Deep Lstm Networks
ICASSP 2018accepted
Controllable generation of emphasis in speech is desirable for expressive TTS systems utilized in various dialog applications. Usually such models remain voice-specific and the strength of emphasis can't be readily controlled. In this work we present a flexible emphatic prosody generation model base…