← Search

Ramya Rasipuram

2 accepted papers

2024

Dialog Modeling in Audiobook Synthesis

ICASSP 2024accepted

In audiobook synthesis, it is important to have the ability to differentiate between dialog and narration or different characters. In this work, we propose dialog modeling methods for audiobook synthesis. The proposed approach consists of two stages. First, a text-based dialog style classifier is em…

Cited by 0SourceScholar
2015

Integrated pronunciation learning for automatic speech recognition using probabilistic lexical modeling

ICASSP 2015accepted

Standard automatic speech recognition (ASR) systems use phoneme-based pronunciation lexicon prepared by linguistic experts. When the hand crafted pronunciations fail to cover the vocabulary of a new domain, a grapheme-to-phoneme (G2P) converter is used to extract pronunciations for new words and the…

Cited by 0SourceScholar