← Search

Minghui Dong

6 accepted papers

2024

A Study on Combining Non-Parallel and Parallel Methodologies for Mandarin-English Cross-Lingual Voice Conversion

ICASSP 2024accepted

In this paper, we propose a cross-lingual voice conversion (VC) scheme leveraging non-parallel and parallel methodologies. The goal of cross-lingual VC is to transform the voice of one speaker from a language dataset into the voice of another speaker from a different language dataset. First, two non…

Cited by 0SourceScholar
2016

A full training framework of cross-stream dependence modelling for HMM-based singing voice synthesis

ICASSP 2016accepted

A cross-stream dependence modelling (CSDM) method has been proposed to model the dependence of spectral distributions on F0 observations for hidden Markov model (HMM) based speech synthesis. However, this method incorporates CSDM only for the embedded training of HMM estimation while ignoring CSDM i…

Cited by 0SourceScholar
2016

Combining multiple kernel models for automatic intelligibility detection of pathological speech

ICASSP 2016accepted

Automatic detection of pathological voice is a challenging task in speech processing. Appropriate acoustic cues of voice can be used to differentiate between normal voices and pathological voices. We propose a method to represent each speech utterance using three types of speech signal representatio…

Cited by 0SourceScholar
2016

Exemplar-based sparse representation of timbre and prosody for voice conversion

ICASSP 2016accepted

Voice conversion (VC) aims to make one speaker (source) to sound like spoken by another speaker (target) without changing the language content. Most of the state-of-the-art voice conversion systems focus only on timbre conversion. However, the speaker identity is characterized by the source-related…

Cited by 0SourceScholar
2015

Sparse representation for frequency warping based voice conversion

ICASSP 2015accepted

This paper presents a sparse representation framework for weighted frequency warping based voice conversion. In this method, a frame-dependent warping function and the corresponding spectral residual vector are first calculated for each source-target spectrum pair. At runtime conversion, a source sp…

Cited by 0SourceScholar