ICASSP 2020accepted0 citations

Learning Asr-Robust Contextualized Embeddings for Spoken Language Understanding

Chao-Wei Huang, Yun-Nung Chen

Abstract

Employing pre-trained language models (LM) to extract contextualized word representations has achieved state-of-the-art performance on various NLP tasks. However, applying this technique to noisy transcripts generated by automatic speech recognizer (ASR) is concerned. Therefore, this paper focuses on making contextualized representations more ASR-robust. We propose a novel confusion-aware fine-tuning method to mitigate the impact of ASR errors on pre-trained LMs. Specifically, we fine-tune LMs to produce similar representations for acoustically confusable words that are obtained from word confusion networks (WCNs) produced by ASR. Experiments on multiple benchmark datasets show that the proposed method significantly improves the performance of spoken language understanding when performing on ASR transcripts.

BibTeX
@inproceedings{icassp2020_learningasrrobus,
  title = {Learning Asr-Robust Contextualized Embeddings for Spoken Language Understanding},
  author = {Chao-Wei Huang and Yun-Nung Chen},
  booktitle = {ICASSP 2020},
  year = {2020}
}
Learning Asr-Robust Contextualized Embeddings for Spoken Language Understanding · ICASSP 2020