2023
SIDGAN: High-Resolution Dubbed Video Generation via Shift-Invariant Learning
ICCV 2023poster
Dubbed video generation aims to accurately synchronize mouth movements of a given facial video with driving audio while preserving identity and scene-specific visual dynamics, such as head pose and lighting. Despite the accurate lip generation of previous approaches that adopts a pretrained audio-vi…