← Search

Ye-Xin Lu

3 accepted papers

2026

DAIEN-TTS: DISENTANGLED AUDIO INFILLING FOR ENVIRONMENT-AWARE TEXT-TO-SPEECH SYNTHESIS

ICASSP 2026poster

This paper presents DAIEN-TTS, a zero-shot text-to-speech (TTS) framework that enables ENvironment-aware synthesis through Disentangled Audio Infilling. By leveraging separate speaker and environment prompts, DAIEN-TTS allows independent control over the timbre and the background environment of the…

Cited by 0SourcePDFScholar
2025

Can Automated Speech Recognition Errors Provide Valuable Clues for Alzheimer's Disease Detection?

ICASSP 2025accepted

Recent advances in automatic speech recognition (ASR) technology have boosted the viability of fully automated Alzheimer’s disease (AD) detection via ASR transcripts. However, there is a lack of understanding of how ASR errors affect the performance of AD detection. This paper addresses that gap. Fi…

Cited by 0SourceScholar
2025

Incremental Disentanglement for Environment-Aware Zero-Shot Text-to-Speech Synthesis

ICASSP 2025accepted

This paper proposes an Incremental Disentanglement-based Environment-Aware zero-shot text-to-speech (TTS) method, dubbed IDEA-TTS, that can synthesize speech for unseen speakers while preserving the acoustic characteristics of a given environment reference speech. IDEA-TTS adopts VITS as the TTS bac…

Cited by 0SourceScholar