ICASSP 2016accepted0 citations

Decoding visemes: Improving machine lip-reading

Helen L. Bear, Richard W. Harvey

Abstract

To undertake machine lip-reading, we try to recognise speech from a visual signal. Current work often uses viseme classification supported by language models with varying degrees of success. A few recent works suggest phoneme classification, in the right circumstances, can outperform viseme classification. In this work we present a novel two-pass method of training phoneme classifiers which uses previously trained visemes in the first pass. With our new training algorithm, we show classification performance which significantly improves on previous lip-reading results.

BibTeX
@inproceedings{icassp2016_decodingvisemesi,
  title = {Decoding visemes: Improving machine lip-reading},
  author = {Helen L. Bear and Richard W. Harvey},
  booktitle = {ICASSP 2016},
  year = {2016}
}
Decoding visemes: Improving machine lip-reading · ICASSP 2016