← Search

Wondong Jang

2 accepted papers

2023

MODEFORMER: Modality-Preserving Embedding For Audio-Video Synchronization Using Transformers

ICASSP 2023accepted

Lack of audio-video synchronization is a common problem during television broadcasts and video conferencing, leading to an unsatisfactory viewing experience. A widely accepted paradigm is to create an error detection mechanism that identifies the cases when audio is leading or lagging. We propose Mo…

Cited by 0SourceScholar
2023

SIDGAN: High-Resolution Dubbed Video Generation via Shift-Invariant Learning

ICCV 2023poster

Dubbed video generation aims to accurately synchronize mouth movements of a given facial video with driving audio while preserving identity and scene-specific visual dynamics, such as head pose and lighting. Despite the accurate lip generation of previous approaches that adopts a pretrained audio-vi…

Cited by 5PDFcodeScholar