2024
Isometric Neural Machine Translation using Phoneme Count Ratio Reward-based Reinforcement Learning
NAACL 2024findings
Traditional Automatic Video Dubbing (AVD) pipeline consists of three key modules, namely, Automatic Speech Recognition (ASR), Neural Machine Translation (NMT), and Text-to-Speech (TTS). Within AVD pipelines, isometric-NMT algorithms are employed to regulate the length of the synthesized output text.…