← Search

Fanbo Meng

3 accepted papers

2022

Improving Adversarial Waveform Generation Based Singing Voice Conversion with Harmonic Signals

ICASSP 2022accepted

Adversarial waveform generation has been a popular approach as the backend of singing voice conversion (SVC) to generate high-quality singing audio. However, the instability of GAN also leads to other problems, such as pitch jitters and U/V errors. It affects the smoothness and continuity of harmoni…

Cited by 0SourceScholar
2021

Inferring Emotion from Large-scale Internet Voice Data: A Semi-supervised Curriculum Augmentation based Deep Learning Approach

AAAI 2021technical

Effective emotion inference from user queries helps to give a more personified response for Voice Dialogue Applications(VDAs). The tremendous amounts of VDA users bring in diverse emotion expressions. How to achieve a high emotion inferring performance from large-scale Internet Voice Data in VDAs? T…

Cited by 16SourcePDFScholar
2015

HMM-based emphatic speech synthesis for corrective feedback in computer-aided pronunciation training

ICASSP 2015accepted

This paper investigates the incorporation of hidden Markov model (HMM) based emphatic speech synthesis for audio exaggeration into an audio-visual speech synthesis framework for the corrective feedback in computer-aided pronunciation training (CAPT). To improve the voice quality of the synthetic emp…

Cited by 0SourceScholar