← Search

Sihun Cha

3 accepted papers

2026

X-AVDT: Audio-Visual Cross-Attention for Robust Deepfake Detection

CVPR 2026

The surge of highly realistic synthetic videos produced by contemporary generative systems has significantly increased the risk of malicious use, challenging both humans and existing detectors. Against this backdrop, we take a generator-side view and observe that internal cross-attention mechanisms

Cited by 0SourceScholar
2025

SALAD: Skeleton-aware Latent Diffusion for Text-driven Motion Generation and Editing

CVPR 2025poster

Text-driven motion generation has advanced significantly with the rise of denoising diffusion models. However, previous methods often oversimplify representations for the skeletal joints, temporal frames, and textual words, limiting their ability to fully capture the information within each modality…

2024

LeGO: Leveraging a Surface Deformation Network for Animatable Stylized Face Generation with One Example

CVPR 2024highlight

Recent advances in 3D face stylization have made significant strides in few to zero-shot settings. However the degree of stylization achieved by existing methods is often not sufficient for practical applications because they are mostly based on statistical 3D Morphable Models (3DMM) with limited va…