2019
Acoustically Grounded Word Embeddings for Improved Acoustics-to-word Speech Recognition
ICASSP 2019accepted
Direct acoustics-to-word (A2W) systems for end-to-end automatic speech recognition are simpler to train, and more efficient to decode with, than sub-word systems. However, A2W systems can have difficulties at training time when data is limited, and at decoding time when recognizing words outside the…