2022
Joint Audio/Text Training for Transformer Rescorer of Streaming Speech Recognition
EMNLP 2022finding
Recently, there has been an increasing interest in two-pass streaming end-to-end speech recognition (ASR) that incorporates a 2nd-pass rescoring model on top of the conventional 1st-pass streaming ASR model to improve recognition accuracy while keeping latency low. One of the latest 2nd-pass rescori…