← Search

Koichiro Ito

3 accepted papers

2021

Audio-Visual Speech Enhancement Method Conditioned in the Lip Motion and Speaker-Discriminative Embeddings

ICASSP 2021accepted

We propose an audio-visual speech enhancement (AVSE) method conditioned both on the speaker’s lip motion and on speaker-discriminative embeddings. We particularly explore a method of extracting the embeddings directly from noisy audio in the AVSE setting without an enrollment procedure. We aim to im…

Cited by 13SourceScholar
2020

Anticipating the Start of User Interaction for Service Robot in the Wild

ICRA 2020poster

A service robot is expected to provide proactive service for visitors who require its help. In contrast to passive service, e.g., providing service only after being spoken to, proactive service initiates an interaction at an early stage, e.g., talking to potential visitors who need the robot’s help…

Cited by 9SourceScholar