← Search

Viet Anh Trinh

5 accepted papers

2024

Automatic Speech Recognition Tuned for Child Speech in the Classroom

ICASSP 2024accepted

K-12 school classrooms have proven to be a challenging environment for Automatic Speech Recognition (ASR) systems, both due to background noise and conversation, and differences in linguistic and acoustic properties from adult speech, on which the majority of ASR systems are trained and evaluated. W…

Cited by 0SourceScholar
2023

Adaptive Endpointing with Deep Contextual Multi-Armed Bandits

ICASSP 2023accepted

Current endpointing (EP) solutions learn in a supervised framework, which does not allow the model to incorporate feedback and improve in an online setting. Also, it is common practice to utilize costly grid-search to find the best configuration for an endpointing model. In this paper, we aim to pro…

Cited by 0SourceScholar
2023

Towards Accurate and Real-Time End-of-Speech Estimation

ICASSP 2023accepted

We introduce a variant of the endpoint (EP) detection problem in automatic speech recognition (ASR), which we call the end-of-speech (EOS) estimation. Given an utterance, EOS estimation aims to identify the timestamp when the utterance waveform has fully decayed and is then used to measure the EP la…

Cited by 0SourceScholar
2022

Unsupervised Speech Enhancement with Speech Recognition Embedding and Disentanglement Losses

ICASSP 2022accepted

Speech enhancement has recently achieved great success with various deep learning methods. However, most conventional speech enhancement systems are trained with supervised methods that impose two significant challenges. First, a majority of training datasets for speech enhancement systems are synth…

Cited by 0SourceScholar