← Search

Leonardo Badino

2 accepted papers

2020

An Analysis of Speech Enhancement and Recognition Losses in Limited Resources Multi-Talker Single Channel Audio-Visual ASR

ICASSP 2020accepted

In this paper, we analyzed how audio-visual speech enhancement can help to perform the ASR task in a cocktail party scenario. Therefore we considered two simple end-to-end LSTM-based models that perform single-channel audiovisual speech enhancement and phone recognition respectively. Then, we studie…

Cited by 0SourceScholar
2019

Face Landmark-based Speaker-independent Audio-visual Speech Enhancement in Multi-talker Environments

ICASSP 2019accepted

In this paper, we address the problem of enhancing the speech of a speaker of interest in a cocktail party scenario when visual information of the speaker of interest is available.Contrary to most previous studies, we do not learn visual features on the typically small audio-visual datasets, but use…

Cited by 0SourceScholar