← Search

George Sung

3 accepted papers

2024

Binaural Angular Separation Network

ICASSP 2024accepted

We propose a neural network model that can separate target speech sources from interfering sources at different angular regions using two microphones. The model is trained with simulated room impulse responses (RIRs) using omnidirectional microphones without needing to collect real RIRs. By relying…

Cited by 0SourceScholar
2024

STREAMVC: Real-Time Low-Latency Voice Conversion

ICASSP 2024accepted

We present StreamVC, a streaming voice conversion solution that preserves the content and prosody of any source speech while matching the voice timbre from any target speech. Unlike previous approaches, StreamVC produces the resulting waveform at low latency from the input signal even on a mobile pl…

Cited by 0SourceScholar
2023

Guided Speech Enhancement Network

ICASSP 2023accepted

High quality speech capture has been widely studied for both voice communication and human computer interface reasons. To improve the capture performance, we can often find multi-microphone speech enhancement techniques deployed on various devices. Multi-microphone speech enhancement problem is ofte…

Cited by 0SourceScholar