← Search

Hanbin Bae

4 accepted papers

2025

Single-Channel Distance-Based Source Separation for Mobile GPU in Outdoor and Indoor Environments

ICASSP 2025accepted

This study emphasizes the significance of exploring distance-based source separation (DSS) in outdoor environments. Unlike existing studies that primarily focus on indoor settings, the proposed model is designed to capture the unique characteristics of outdoor audio sources. It incorporates advanced…

Cited by 0SourceScholar
2024

FINALLY: fast and universal speech enhancement with studio-like quality

NeurIPS 2024poster

In this paper, we address the challenge of speech enhancement in real-world recordings, which often contain various forms of distortion, such as background noise, reverberation, and microphone artifacts. We revisit the use of Generative Adversarial Networks (GANs) for speech enhancement and theoreti…

2023

Avocodo: Generative Adversarial Network for Artifact-Free Vocoder

AAAI 2023technical

Neural vocoders based on the generative adversarial neural network (GAN) have been widely used due to their fast inference speed and lightweight networks while generating high-quality speech waveforms. Since the perceptually important speech components are primarily concentrated in the low-frequency…

2021

A Neural Text-to-Speech Model Utilizing Broadcast Data Mixed with Background Music

ICASSP 2021accepted

Recently, it has become easier to obtain speech data from various media such as the internet or YouTube, but directly utilizing them to train a neural text-to-speech (TTS) model is difficult. The proportion of clean speech is insufficient and the remainder includes background music. Even with the gl…

Cited by 0SourceScholar