← Search

Parnia Bahar

3 accepted papers

2025

Maintaining Prosodic Consistency in Automatic Dubbing for Better Isochrony

ICASSP 2025accepted

One of the major requirements of automatic dubbing, as an extension of speech-to-speech translation, is isochrony. It refers to the fact that dubbed speech should line up in time with the original speech, matching the phrase-pause arrangement of the source utterance. To achieve this, we explore and…

Cited by 0SourceScholar
2020

Exploring A Zero-Order Direct Hmm Based on Latent Attention for Automatic Speech Recognition

ICASSP 2020accepted

In this paper, we study a simple yet elegant latent variable attention model for automatic speech recognition (ASR) which enables an integration of attention sequence modeling into the direct hidden Markov model (HMM) concept. We use a sequence of hidden variables that establishes a mapping from out…

Cited by 0SourceScholar
2019

On Using 2D Sequence-to-sequence Models for Speech Recognition

ICASSP 2019accepted

Attention-based sequence-to-sequence models have shown promising results in automatic speech recognition. Using these architectures, one-dimensional input and output sequences are related by an attention approach, thereby replacing more explicit alignment processes, like in classical HMM-based model…

Cited by 0SourceScholar