← Search

Giovanni Morrone

3 accepted papers

2021

Audio-Visual Speech Inpainting with Deep Learning

ICASSP 2021accepted

In this paper, we present a deep-learning-based framework for audio-visual speech inpainting, i.e., the task of restoring the missing parts of an acoustic speech signal from reliable audio context and uncorrupted visual information. Recent work focuses solely on audio-only methods and generally aims…

Cited by 31SourceScholar
2020

An Analysis of Speech Enhancement and Recognition Losses in Limited Resources Multi-Talker Single Channel Audio-Visual ASR

ICASSP 2020accepted

In this paper, we analyzed how audio-visual speech enhancement can help to perform the ASR task in a cocktail party scenario. Therefore we considered two simple end-to-end LSTM-based models that perform single-channel audiovisual speech enhancement and phone recognition respectively. Then, we studie…

Cited by 3SourceScholar
2019

Face Landmark-based Speaker-independent Audio-visual Speech Enhancement in Multi-talker Environments

ICASSP 2019accepted

In this paper, we address the problem of enhancing the speech of a speaker of interest in a cocktail party scenario when visual information of the speaker of interest is available.Contrary to most previous studies, we do not learn visual features on the typically small audio-visual datasets, but use…

Cited by 0SourceScholar