← Search

Justin Chan

2 accepted papers

2023

Real-Time Target Sound Extraction

ICASSP 2023accepted

We present the first neural network model to achieve real-time and streaming target sound extraction. To accomplish this, we propose Waveformer, an encoder-decoder architecture with a stack of dilated causal convolution layers as the encoder, and a transformer decoder layer as the decoder. This hybr…

Cited by 0SourceScholar