2019
Joint Endpointing and Decoding with End-to-end Models
ICASSP 2019accepted
The tradeoff between word error rate (WER) and latency is very important for streaming automatic speech recognition (ASR) applications. We want the system to endpoint and close the microphone as quickly as possible, without degrading WER. Conventional ASR systems rely on a separately trained endpoin…