ICASSP 2018accepted0 citations
Adaptation of an Expressive Single Speaker Deep Neural Network Speech Synthesis System
Jonathan Parker, Yannis Stylianou, Roberto Cipolla
Abstract
One of the advantages of statistical parametric speech synthesis is the ability to alter some of the characteristics of the speech e.g. change the speaker, expression etc. In this paper we present a technique to adapt an expressive single speaker deep neural network (DNN) speech synthesis model to a new speaker, allowing for both neutral and expressive speech in the new speaker's voice. Experiments show that the proposed adaptation technique achieves higher MOS scores on both neutral and expressive speech, and higher speaker similarity and slightly lower expression similarity scores on the expressive speech when compared with another DNN speaker adaptation technique.
BibTeX
@inproceedings{icassp2018_adaptationofanex,
title = {Adaptation of an Expressive Single Speaker Deep Neural Network Speech Synthesis System},
author = {Jonathan Parker and Yannis Stylianou and Roberto Cipolla},
booktitle = {ICASSP 2018},
year = {2018}
}