2022
Learning to Predict Speech in Silent Videos Via Audiovisual Analogy
ICASSP 2022accepted
Lipreading is a difficult task, even for humans. And synthesizing the original speech waveform from lipreading makes it even a more challenging problem. Towards this end, we present a deep learning framework that can be trained end-to-end to learn the mapping between the auditory and visual signals.…