← Search

Hao Che

3 accepted papers

2023

Improving Prosody for Cross-Speaker Style Transfer by Semi-Supervised Style Extractor and Hierarchical Modeling in Speech Synthesis

ICASSP 2023accepted

Cross-speaker style transfer in speech synthesis aims at transferring a style from source speaker to synthesized speech of a target speaker’s timbre. In most previous methods, the synthesized fine-grained prosody features often represent the source speaker’s average style, similar to the one-to-many…

Cited by 0SourceScholar
2022

K-Converter: An Unsupervised Singing Voice Conversion System

ICASSP 2022accepted

Singing voice conversion (SVC) converts a singer’s voice to another one’s voice while preserving the linguistic content. Recently, some SVC systems rely on supervised phonetic features extracted from pre-trained automatic speech recognition (ASR) models, increasing system complexity. Some end-toend…

Cited by 0SourceScholar
2021

One-Shot Voice Conversion Based on Speaker Aware Module

ICASSP 2021accepted

Voice conversion (VC) is a task to convert the voice of speech while preserving its linguistic content. Although several methods have been proposed to enable VC with non-parallel data, it is still difficult to model the voice without a great number of data or an adaptive process. In this paper, we p…

Cited by 0SourceScholar