← Search

Weipeng He

5 accepted papers

2025

Directional Source Separation for Robust Speech Recognition on Smart Glasses

ICASSP 2025accepted

Modern smart glasses leverage machine learning to offer real-time transcriptions, considerably enriching human communication experiences. However, such systems frequently encounter challenges related to environmental noises, leading to decreased speech recognition. To improve voice quality, this wor…

Cited by 15SourceScholar
2023

Egocentric Audio-Visual Noise Suppression

ICASSP 2023accepted

This paper studies audio-visual noise suppression for egocentric videos -where the speaker is not captured in the video. Instead, potential noise sources are visible on screen with the camera emulating the off-screen speaker’s view of the outside world. This setting is different from prior work in a…

Cited by 0SourceScholar
2020

Spatial Attention for Far-Field Speech Recognition with Deep Beamforming Neural Networks

ICASSP 2020accepted

In this paper, we introduce spatial attention for refining the information in multi-direction neural beamformer for far-field automatic speech recognition. Previous approaches of neural beamformers with multiple look directions, such as the factored complex linear projection, have shown promising re…

Cited by 0SourceScholar
2019

Adaptation of Multiple Sound Source Localization Neural Networks with Weak Supervision and Domain-adversarial Training

ICASSP 2019accepted

Despite the recent success of deep neural network-based approaches in sound source localization, these approaches suffer the limitations that the required annotation process is costly, and the mismatch between the training and test conditions undermines the performance. This paper addresses the ques…

Cited by 0SourceScholar