← Search

Robert A. J. Clark

3 accepted papers

2016

Deep neural network-guided unit selection synthesis

ICASSP 2016accepted

Vocoding of speech is a standard part of statistical parametric speech synthesis systems. It imposes an upper bound of the naturalness that can possibly be achieved. Hybrid systems using parametric models to guide the selection of natural speech units can combine the benefits of robust statistical m…

Cited by 0SourceScholar
2016

Wavelet-based decomposition of F0 as a secondary task for DNN-based speech synthesis with multi-task learning

ICASSP 2016accepted

We investigate two wavelet-based decomposition strategies of the f0 signal and their usefulness as a secondary task for speech synthesis using multi-task deep neural networks (MTL-DNN). The first decomposition strategy uses a static set of scales for all utterances in the training data. We propose a…

Cited by 12SourceScholar
2015

A multi-level representation of f0 using the continuous wavelet transform and the Discrete Cosine Transform

ICASSP 2015accepted

We propose a representation of f0 using the Continuous Wavelet Transform (CWT) and the Discrete Cosine Transform (DCT). The CWT decomposes the signal into various scales of selected frequencies, while the DCT compactly represents complex contours as a weighted sum of cosine functions. The proposed a…

Cited by 0SourceScholar