← Search

Sergio Elizondo

1 accepted papers

2023

SIDGAN: High-Resolution Dubbed Video Generation via Shift-Invariant Learning

ICCV 2023poster

Dubbed video generation aims to accurately synchronize mouth movements of a given facial video with driving audio while preserving identity and scene-specific visual dynamics, such as head pose and lighting. Despite the accurate lip generation of previous approaches that adopts a pretrained audio-vi…

Cited by 5PDFcodeScholar