← Search

Yi Meng

2 accepted papers

2022

Enhancing Speaking Styles in Conversational Text-to-Speech Synthesis with Graph-Based Multi-Modal Context Modeling

ICASSP 2022accepted

Comparing with traditional text-to-speech (TTS) systems, conversational TTS systems are required to synthesize speeches with proper speaking style confirming to the conversational context. However, state-of-the-art context modeling methods in conversational TTS only model the textual information in…

Cited by 0SourceScholar
2022

Neufa: Neural Network Based End-to-End Forced Alignment with Bidirectional Attention Mechanism

ICASSP 2022accepted

Although deep learning and end-to-end models have been widely used and shown their superiority in automatic speech recognition (ASR) and text-to-speech (TTS) synthesis, state-of-the-art forced alignment (FA) models are still based on hidden Markov model (HMM). HMM has limited view of contextual info…

Cited by 0SourceScholar