← Search

Hyun-Wook Yoon

2 accepted papers

2024

Enhancing Multilingual TTS with Voice Conversion Based Data Augmentation and Posterior Embedding

ICASSP 2024accepted

This paper proposes a multilingual, multi-speaker (MM) TTS system by using a voice conversion (VC)-based data augmentation method. Creating an MM-TTS model is challenging, owing to the difficulties of collecting polyglot data from multiple speakers. To address this problem, we adopt a cross-lingual,…

Cited by 0SourceScholar
2021

Multi-SpectroGAN: High-Diversity and High-Fidelity Spectrogram Generation with Adversarial Style Combination for Speech Synthesis

AAAI 2021technical

While generative adversarial networks (GANs) based neural text-to-speech (TTS) systems have shown significant improvement in neural speech synthesis, there is no TTS system to learn to synthesize speech from text sequences with only adversarial feedback. Because adversarial feedback alone is not suf…

Cited by 65SourcePDFScholar