ICASSP 2017accepted0 citations

Sequence-to-sequence models for punctuated transcription combining lexical and acoustic features

Ondrej Klejch, Peter Bell, Steve Renals

Abstract

In this paper we present an extension of our previously described neural machine translation based system for punctuated transcription. This extension allows the system to map from per frame acoustic features to word level representations by replacing the traditional encoder in the encoder-decoder architecture with a hierarchical encoder. Furthermore, we show that a system combining lexical and acoustic features significantly outperforms systems using only a single source of features on all measured punctuation marks. The combination of lexical and acoustic features achieves a significant improvement in F-Measure of 1.5 absolute over the purely lexical neural machine translation based system.

BibTeX
@inproceedings{icassp2017_sequencetosequen,
  title = {Sequence-to-sequence models for punctuated transcription combining lexical and acoustic features},
  author = {Ondrej Klejch and Peter Bell and Steve Renals},
  booktitle = {ICASSP 2017},
  year = {2017}
}