← Search

Jinhyeok Yang

3 accepted papers

2023

Avocodo: Generative Adversarial Network for Artifact-Free Vocoder

AAAI 2023technical

Neural vocoders based on the generative adversarial neural network (GAN) have been widely used due to their fast inference speed and lightweight networks while generating high-quality speech waveforms. Since the perceptually important speech components are primarily concentrated in the low-frequency…

2023

NANSY++: Unified Voice Synthesis with Neural Analysis and Synthesis

ICLR 2023poster

Various applications of voice synthesis have been developed independently despite the fact that they generate “voice” as output in common. In addition, most of the voice synthesis models still require a large number of audio data paired with annotated labels (e.g., text transcription and music score…

Cited by 62SourcePDFScholar
2022

Varianceflow: High-Quality and Controllable Text-to-Speech using Variance Information via Normalizing Flow

ICASSP 2022accepted

There are two types of methods for non-autoregressive text-to-speech models to learn the one-to-many relationship between text and speech effectively. The first one is to use an advanced generative framework such as normalizing flow (NF). The second one is to use variance information such as pitch o…

Cited by 0SourceScholar