← Search

Tyler Vuong

2 accepted papers

2023

Unsupervised Voice Type Discrimination Score Adaptation Using X-Vector Clusters

ICASSP 2023accepted

Voice type discrimination (VTD) is the task of automatically detecting speech produced in the same room as a recording device ("live speech") among other speech and non-speech noises, such as traffic noises or radio broadcasts ("distractor audio"). Existing work has described methods for performing…

Cited by 0SourceScholar
2021

A Modulation-Domain Loss for Neural-Network-Based Real-Time Speech Enhancement

ICASSP 2021accepted

We describe a modulation-domain loss function for deep-learning-based speech enhancement systems. Learnable spectro-temporal receptive fields (STRFs) were adapted to optimize for a speaker identification task. The learned STRFs were then used to calculate a weighted mean-squared error (MSE) in the m…

Cited by 0SourceScholar