← Search

Longhao Zhang

4 accepted papers

2025

DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance

ICCV 2025poster

While recent image-based human animation methods achieve realistic body and facial motion synthesis, critical gaps remain in fine-grained holistic controllability, multi-scale adaptability, and long-term temporal coherence, which leads to their lower expressiveness and robustness. We propose a diffu…

Cited by 0SourcePDFScholar
2025

INFP: Audio-Driven Interactive Head Generation in Dyadic Conversations

CVPR 2025poster

Imagine having a conversation with a socially intelligent agent. It can attentively listen to your words and offer visual and linguistic feedback promptly. This seamless interaction allows for multiple rounds of conversation to flow smoothly and naturally. In pursuit of actualizing it, we propose IN…

Cited by 4SourcePDFScholar
2023

One-Shot High-Fidelity Talking-Head Synthesis With Deformable Neural Radiance Field

CVPR 2023poster

Talking head generation aims to generate faces that maintain the identity information of the source image and imitate the motion of the driving image. Most pioneering methods rely primarily on 2D representations and thus will inevitably suffer from face distortion when large head rotations are encou…

Cited by 54SourcePDFScholar
2022

Depth-Aware Generative Adversarial Network for Talking Head Video Generation

CVPR 2022poster

Talking head video generation aims to produce a synthetic human face video that contains the identity and pose information respectively from a given source image and a driving video. Existing works for this task heavily rely on 2D representations (e.g. appearance and motion) learned from the input i…

Cited by 203PDFcodeScholar