← Search

Gangjian Zhang

5 accepted papers

2026

FastAnimate: Towards Learnable Template Construction and Pose Deformation for Fast 3D Human Avatar Animation

AAAI 2026technical

3D human avatar animation aims at transforming a human avatar from an arbitrary initial pose to a specified target pose using deformation algorithms. Existing approaches typically divide this task into two stages: canonical template construction and target pose deformation. However, current template

Cited by 0SourcePDFScholar
2025

Diversified Augmentation with Domain Adaptation for Debiased Video Temporal Grounding

ICASSP 2025accepted

Temporal sentence grounding in videos (TSGV) faces challenges due to public TSGV datasets containing significant temporal biases, which are attributed to the uneven temporal distributions of target moments. Existing methods generate augmented videos, where target moments are forced to have varying t…

Cited by 0SourceScholar
2025

DoGA: Enhancing Grounded Object Detection via Grouped Pre-Training with Attributes

AAAI 2025technical

Recent advances in vision-language pre-training have significantly enhanced the model capabilities on grounded object detection. However, these studies often pre-train with coarse-grained text prompts, such as plain category names and brief grounded phrases. This limitation curtails the model's capa…

2025

Graph-Guided Scene Reconstruction from Images with 3D Gaussian Splatting

ICLR 2025poster

This paper investigates an open research challenge of reconstructing high-quality, large-scale 3D open scenes from images. It is observed existing methods have various limitations, such as requiring precise camera poses for input and dense viewpoints for supervision. To perform effective and effici…

Cited by 1SourcePDFScholar
2025

MultiGO: Towards Multi-level Geometry Learning for Monocular 3D Textured Human Reconstruction

CVPR 2025poster

This paper investigates the research task of reconstructing the 3D clothed human body from a monocular image. Due to the inherent ambiguity of single-view input, existing approaches leverage pre-trained SMPL(-X) estimation models or generative models to provide auxiliary information for human recons…

Cited by 3SourcePDFScholar