← Search

Runxuan Yang

3 accepted papers

2025

SonicSim: A customizable simulation platform for speech processing in moving sound source scenarios

ICLR 2025poster

Systematic evaluation of speech separation and enhancement models under moving sound source conditions requires extensive and diverse data. However, real-world datasets often lack sufficient data for training and evaluation, and synthetic datasets, while larger, lack acoustic realism. Consequently,…

2024

IIANet: An Intra- and Inter-Modality Attention Network for Audio-Visual Speech Separation

ICML 2024poster

Recent research has made significant progress in designing fusion modules for audio-visual speech separation. However, they predominantly focus on multi-modal fusion at a single temporal scale of auditory and visual features without employing selective attention mechanisms, which is in sharp contras…

2023

An efficient encoder-decoder architecture with top-down attention for speech separation

ICLR 2023poster

Deep neural networks have shown excellent prospects in speech separation tasks. However, obtaining good results while keeping a low model complexity remains challenging in real-world applications. In this paper, we provide a bio-inspired efficient encoder-decoder architecture by mimicking the brain’…