← Search

Thi Ngoc Tho Nguyen

9 accepted papers

2022

End-to-End Complex-Valued Multidilated Convolutional Neural Network for Joint Acoustic Echo Cancellation and Noise Suppression

ICASSP 2022accepted

Echo and noise suppression is an integral part of a full-duplex communication system. Many recent acoustic echo cancellation (AEC) systems rely on a separate adaptive filtering module for linear echo suppression and a neural module for residual echo suppression. However, in practice, adaptive filter…

Cited by 0SourceScholar
2022

Polyphonic Audio Event Detection: Multi-Label or Multi-Class Multi-Task Classification Problem?

ICASSP 2022accepted

Polyphonic events are the main error source of audio event detection (AED) systems. In deep-learning context, the most common approach to deal with event overlaps is to treat the AED task as a multi-label classification problem. By doing this, we inherently consider multiple one-vs.-rest classificat…

Cited by 0SourceScholar
2022

SALSA-Lite: A Fast and Effective Feature for Polyphonic Sound Event Localization and Detection with Microphone Arrays

ICASSP 2022accepted

Polyphonic sound event localization and detection (SELD) has many practical applications in acoustic sensing and monitoring. However, the development of real-time SELD has been limited by the demanding computational requirement of most recent SELD systems. In this work, we introduce SALSA-Lite, a fa…

Cited by 0SourceScholar
2021

A General Network Architecture for Sound Event Localization and Detection Using Transfer Learning and Recurrent Neural Network

ICASSP 2021accepted

Polyphonic sound event detection and localization (SELD) task is challenging because it is difficult to jointly optimize sound event detection (SED) and direction-of-arrival (DOA) estimation in the same network. We propose a general network architecture for SELD in which the SELD network comprises s…

Cited by 0SourceScholar
2020

A Sequence Matching Network for Polyphonic Sound Event Localization and Detection

ICASSP 2020accepted

Polyphonic sound event detection and direction-of-arrival estimation require different input features from audio signals. While sound event detection mainly relies on time-frequency patterns, direction-of-arrival estimation relies on magnitude or phase differences between microphones. Previous appro…

Cited by 0SourceScholar
2017

A novel sparse model for multi-source localization using distributed microphone array

ICASSP 2017accepted

When distances between microphone pairs are larger than the half-wavelength of signals, source localization methods using cross-correlation such as time-difference-of-arrival (TDOA), steered response power (SRP) are commonly used in practice. We present here a novel model that expresses microphone p…

Cited by 0SourceScholar
2016

An expectation-maximization eigenvector clustering approach to direction of arrival estimation of multiple speech sources

ICASSP 2016accepted

This paper presents an eigenvector clustering approach for estimating the direction of arrival (DOA) of multiple speech signals using a microphone array. Existing clustering approaches usually only use low frequencies to avoid spatial aliasing. In this study, we propose a probabilistic eigenvector c…

Cited by 0SourceScholar
2016

Large region acoustic source mapping: A generalized sparse constrained deconvolution approach

ICASSP 2016accepted

This paper presents a generalized multiple-point sparse constrained deconvolution approach for mapping acoustic noise sources in large regions using a movable array. Extended from our previous MPSC-DAMAS approach, we first derive a generalized inverse problem relating to the source powers and the ar…

Cited by 0SourceScholar