← Search

Junchen Lu

2 accepted papers

2026

PERFORMSINGER: MULTIMODAL SINGING VOICE SYNTHESIS LEVERAGING SYNCHRONIZED LIP CUES FROM SINGING PERFORMANCE VIDEOS

ICASSP 2026poster

Existing singing voice synthesis (SVS) models largely rely on fine-grained, phoneme-level durations, which limits their practical application. These methods overlook the complementary role of visual information in duration prediction.To address these issues, we propose PerformSinger, a pioneering mu…

Cited by 0SourcePDFScholar
2022

Visualtts: TTS with Accurate Lip-Speech Synchronization for Automatic Voice Over

ICASSP 2022accepted

In this paper, we formulate a novel task to synthesize speech in sync with a silent pre-recorded video, denoted as automatic voice over (AVO). Unlike traditional speech synthesis, AVO seeks to generate not only human-sounding speech, but also perfect lip-speech synchronization. A natural solution to…

Cited by 0SourceScholar