2021
Multi-Speaker Emotional Speech Synthesis with Fine-Grained Prosody Modeling
ICASSP 2021accepted
We present an end-to-end system for multi-speaker emotional speech synthesis. In particular, our system learns emotion classes from just two speakers then generalizes these classes to other speakers from whom no emotional data was seen. We address the problem by integrating disentangled, fine-graine…