← Search

Juheon Lee

4 accepted papers

2023

NANSY++: Unified Voice Synthesis with Neural Analysis and Synthesis

ICLR 2023poster

Various applications of voice synthesis have been developed independently despite the fact that they generate “voice” as output in common. In addition, most of the voice synthesis models still require a large number of audio data paired with annotated labels (e.g., text transcription and music score…

Cited by 62SourcePDFScholar
2021

Neural Analysis and Synthesis: Reconstructing Speech from Self-Supervised Representations

NeurIPS 2021poster

We present a neural analysis and synthesis (NANSY) framework that can manipulate the voice, pitch, and speed of an arbitrary speech signal. Most of the previous works have focused on using information bottleneck to disentangle analysis features for controllable synthesis, which usually results in p…

Cited by 178SourcePDFScholar
2020

Disentangling Timbre and Singing Style with Multi-Singer Singing Synthesis System

ICASSP 2020accepted

In this study, we define the identity of the singer with two independent concepts – timbre and singing style – and propose a multi-singer singing synthesis system that can model them separately. To this end, we extend our single-singer model into a multi-singer model in the following ways: first, we…

Cited by 0SourceScholar
2018

Cover Song Identification Using Song-to-Song Cross-Similarity Matrix with Convolutional Neural Network

ICASSP 2018accepted

In this paper, we propose a cover song identification algorithm using a convolutional neural network (CNN). We first train the CNN model to classify any non-/cover relationship, by feeding a cross-similarity matrix that is generated from a pair of songs as an input. Our main idea is to use the CNN o…

Cited by 0SourceScholar