← Search

Jingkai Zhou

8 accepted papers

2026

RealisMotion: Decomposed Human Motion Control and Video Generation in the World Space

ICML 2026poster

Generating human videos with realistic and controllable motions is a challenging task. While existing methods can generate visually compelling videos, they lack separate control over four key video elements: foreground subject, background video, human trajectory, and action patterns. In this paper, …

Cited by 0SourceScholar
2025

Layer-Animate for Transparent Video Generation

ICASSP 2025accepted

Transparent videos with alpha channels play a crucial role in film production, advertising, and augmented reality fields. However, there is currently no available method for producing transparent videos. Traditional methods are time-consuming and labor-intensive, and employing alternative approaches…

Cited by 0SourceScholar
2025

On Denoising Walking Videos for Gait Recognition

CVPR 2025poster

To capture individual gait patterns, excluding identity-irrelevant cues in walking videos, such as clothing texture and color, remains a persistent challenge for vision-based gait recognition. Traditional silhouette and pose-based methods, though theoretically effective at removing such distractions…

2025

RealisHuman: A Two-Stage Approach for Refining Malformed Human Parts in Generated Images

AAAI 2025technical

In recent years, diffusion models have revolutionized visual generation, outperforming traditional frameworks like Generative Adversarial Networks (GANs). However, generating images of humans with realistic semantic parts, such as hands and faces, remains a significant challenge due to their intric…

2024

Adversarial Score Distillation: When score distillation meets GAN

CVPR 2024poster

Existing score distillation methods are sensitive to classifier-free guidance (CFG) scale manifested as over-smoothness or instability at small CFG scales while over-saturation at large ones. To explain and analyze these issues we revisit the derivation of Score Distillation Sampling (SDS) and decip…

2023

ASM: Adaptive Skinning Model for High-Quality 3D Face Modeling

ICCV 2023poster

The research fields of parametric face model and 3D face reconstruction have been extensively studied. However, a critical question remains unanswered: how to tailor the face model for specific reconstruction settings. We argue that reconstruction with multi-view uncalibrated images demands a new mo…

Cited by 6PDFScholar
2022

Scaled ReLU Matters for Training Vision Transformers

AAAI 2022technical

Vision transformers (ViTs) have been an alternative design paradigm to convolutional neural networks (CNNs). However, the training of ViTs is much harder than CNNs, as it is sensitive to the training parameters, such as learning rate, optimizer and warmup epoch. The reasons for training difficulty a…

Cited by 46SourcePDFScholar