← Search

Tuomo Raitio

3 accepted papers

2024

Dialog Modeling in Audiobook Synthesis

ICASSP 2024accepted

In audiobook synthesis, it is important to have the ability to differentiate between dialog and narration or different characters. In this work, we propose dialog modeling methods for audiobook synthesis. The proposed approach consists of two stages. First, a text-based dialog style classifier is em…

Cited by 0SourceScholar
2022

Hierarchical Prosody Modeling and Control in Non-Autoregressive Parallel Neural TTS

ICASSP 2022accepted

Neural text-to-speech (TTS) synthesis can generate speech that is indistinguishable from natural speech. However, the synthetic speech often represents the average prosodic style of the database instead of having more versatile prosodic variation. Moreover, many models lack the ability to control th…

Cited by 0SourceScholar