2017
Feature extraction using multimodal convolutional neural networks for visual speech recognition
ICASSP 2017accepted
This article addresses the problem of continuous speech recognition from visual information only, without exploiting any audio signal. Our approach combines a video camera and an ultrasound imaging system for monitoring simultaneously the speaker's lips and the movement of the tongue. We investigate…