← Search

Ashish Sardana

2 accepted papers

2022

Learning to Predict Speech in Silent Videos Via Audiovisual Analogy

ICASSP 2022accepted

Lipreading is a difficult task, even for humans. And synthesizing the original speech waveform from lipreading makes it even a more challenging problem. Towards this end, we present a deep learning framework that can be trained end-to-end to learn the mapping between the auditory and visual signals.…

Cited by 0SourceScholar
2021

Speech Prediction in Silent Videos Using Variational Autoencoders

ICASSP 2021accepted

Understanding the relationship between the auditory and visual signals is crucial for many different applications ranging from computer-generated imagery (CGI) and video editing automation to assisting people with hearing or visual impairments. However, this is challenging since the distribution of…

Cited by 0SourceScholar