← Search

Ravindra Yadav

3 accepted papers

2022

Emotion-Controllable Generalized Talking Face Generation

IJCAI 2022poster

Despite the significant progress in recent years, very few of the AI-based talking face generation methods attempt to render natural emotions. Moreover, the scope of the methods is majorly limited to the characteristics of the training dataset, hence they fail to generalize to arbitrary unseen faces…

2022

Learning to Predict Speech in Silent Videos Via Audiovisual Analogy

ICASSP 2022accepted

Lipreading is a difficult task, even for humans. And synthesizing the original speech waveform from lipreading makes it even a more challenging problem. Towards this end, we present a deep learning framework that can be trained end-to-end to learn the mapping between the auditory and visual signals.…

Cited by 0SourceScholar
2021

Speech Prediction in Silent Videos Using Variational Autoencoders

ICASSP 2021accepted

Understanding the relationship between the auditory and visual signals is crucial for many different applications ranging from computer-generated imagery (CGI) and video editing automation to assisting people with hearing or visual impairments. However, this is challenging since the distribution of…

Cited by 0SourceScholar