← Search

Sibo Zhang

2 accepted papers

2022

Text2video: Text-Driven Talking-Head Video Synthesis with Personalized Phoneme - Pose Dictionary

ICASSP 2022accepted

With the advance of deep learning technology, automatic video generation from audio or text has become an emerging and promising research topic. In this paper, we present a novel approach to synthesize video from the text. The method builds a phoneme-pose dictionary and trains a generative adversari…

Cited by 0SourceScholar
2020

DVI: Depth Guided Video Inpainting for Autonomous Driving

ECCV 2020poster

To get clear street-view and photo-realistic simulation in autonomous driving, we present an automatic video inpainting algorithm that can remove traffic agents from videos and synthesize missing regions with the guidance of depth/point cloud. By building a dense 3D map from stitched point clouds, f…