ICASSP 2018accepted0 citations

Adaptation of an Expressive Single Speaker Deep Neural Network Speech Synthesis System

Jonathan Parker, Yannis Stylianou, Roberto Cipolla

Abstract

One of the advantages of statistical parametric speech synthesis is the ability to alter some of the characteristics of the speech e.g. change the speaker, expression etc. In this paper we present a technique to adapt an expressive single speaker deep neural network (DNN) speech synthesis model to a new speaker, allowing for both neutral and expressive speech in the new speaker's voice. Experiments show that the proposed adaptation technique achieves higher MOS scores on both neutral and expressive speech, and higher speaker similarity and slightly lower expression similarity scores on the expressive speech when compared with another DNN speaker adaptation technique.

BibTeX
@inproceedings{icassp2018_adaptationofanex,
  title = {Adaptation of an Expressive Single Speaker Deep Neural Network Speech Synthesis System},
  author = {Jonathan Parker and Yannis Stylianou and Roberto Cipolla},
  booktitle = {ICASSP 2018},
  year = {2018}
}