← Search

Xucheng Wan

3 accepted papers

2025

CAMEL: Cross-Attention Enhanced Mixture-of-Experts and Language Bias for Code-Switching Speech Recognition

ICASSP 2025accepted

Code-switching automatic speech recognition (ASR) aims to transcribe speech that contains two or more languages accurately. To better capture language-specific speech representations and address language confusion in code-switching ASR, the mixture-of-experts (MoE) architecture and an additional lan…

Cited by 0SourceScholar
2025

SCDiar: a streaming diarization system based on speaker change detection and speech recognition

ICASSP 2025accepted

In hours-long meeting scenarios, real-time speech stream often struggles with achieving accurate speaker diarization, commonly leading to speaker identification and speaker count errors. To address this challenge, we propose SCDiar, a system that operates on speech segments, split at the token level…

Cited by 0SourceScholar
2023

X-SEPFORMER: End-To-End Speaker Extraction Network with Explicit Optimization on Speaker Confusion

ICASSP 2023accepted

Target speech extraction (TSE) systems are designed to extract target speech from a multi-talker mixture. The popular training objective for most prior TSE networks is to enhance reconstruction performance of extracted speech waveform. However, it has been reported that a TSE system delivers high re…

Cited by 0SourceScholar