← Search

Ricong Huang

2 accepted papers

2024

VersVideo: Leveraging Enhanced Temporal Diffusion Models for Versatile Video Generation

ICLR 2024poster

Creating stable, controllable videos is a complex task due to the need for significant variation in temporal dynamics and cross-frame temporal consistency. To address this, we enhance the spatial-temporal capability and introduce a versatile video generation model, VersVideo, which leverages textual…

2023

Parametric Implicit Face Representation for Audio-Driven Facial Reenactment

CVPR 2023poster

Audio-driven facial reenactment is a crucial technique that has a range of applications in film-making, virtual avatars and video conferences. Existing works either employ explicit intermediate face representations (e.g., 2D facial landmarks or 3D face models) or implicit ones (e.g., Neural Radiance…

Cited by 20SourcePDFScholar