← Search

Yulan Liu

4 accepted papers

2022

Multi-Turn RNN-T for Streaming Recognition of Multi-Party Speech

ICASSP 2022accepted

Automatic speech recognition (ASR) of single channel far-field recordings with an unknown number of speakers is traditionally tackled by cascaded modules. Recent research shows that end-to-end (E2E) multi-speaker ASR models can achieve superior recognition accuracy compared to modular systems. Howev…

Cited by 0SourceScholar
2021

Using Synthetic Audio to Improve the Recognition of Out-of-Vocabulary Words in End-to-End Asr Systems

ICASSP 2021accepted

Today, many state-of-the-art automatic speech recognition (ASR) systems apply all-neural models that map audio to word sequences trained end-to-end along one global optimisation criterion in a fully data driven fashion. These models allow high precision ASR for domains and words represented in the t…

Cited by 0SourceScholar