← Search

Jiaran Cai

2 accepted papers

2026

Training-Free Multi-Character Audio-Driven Animation via Diffusion Transformer with Reward Feedback

AAAI 2026technical

Recent advances in diffusion models have significantly improved audio-driven human video generation, surpassing traditional methods in both quality and controllability. However, existing approaches still face challenges in lip-sync accuracy, temporal coherence for long video generation, and multi-ch

Cited by 0SourcePDFScholar
2025

Playmate: Flexible Control of Portrait Animation via 3D-Implicit Space Guided Diffusion

ICML 2025poster

Recent diffusion-based talking face generation models have demonstrated impressive potential in synthesizing videos that accurately match a speech audio clip with a given reference identity. However, existing approaches still encounter significant challenges due to uncontrollable factors, such as in…