← Search

Jianqiang Ren

6 accepted papers

2026

Mitigating Error Accumulation in Co-Speech Motion Generation via Global Rotation Diffusion and Multi-Level Constraints

AAAI 2026technical

Reliable co-speech motion generation requires precise motion representation and consistent structural priors across all joints. Existing generative methods typically operate on local joint rotations, which are defined hierarchically based on the skeleton structure. This leads to cumulative errors du

Cited by 0SourcePDFScholar
2025

SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis

ICCV 2025poster

A good co-speech motion generation cannot be achieved without a careful integration of common rhythmic motion and rare yet essential semantic motion. In this work, we propose SemTalk for holistic co-speech motion generation with frame-level semantic emphasis. Our key insight is to separately learn b…

Cited by 0SourcePDFScholar
2023

A Hierarchical Representation Network for Accurate and Detailed Face Reconstruction From In-the-Wild Images

CVPR 2023poster

Limited by the nature of the low-dimensional representational capacity of 3DMM, most of the 3DMM-based face reconstruction (FR) methods fail to recover high-frequency facial details, such as wrinkles, dimples, etc. Some attempt to solve the problem by introducing detail maps or non-linear operations…

2022

Structure-Aware Flow Generation for Human Body Reshaping

CVPR 2022poster

Body reshaping is an important procedure in portrait photo retouching. Due to the complicated structure and multifarious appearance of human bodies, existing methods either fall back on the 3D domain via body morphable model or resort to keypoint-based image deformation, leading to inefficiency and…

Cited by 7PDFcodeScholar
2019

Attention-Aware Multi-Stroke Style Transfer

CVPR 2019poster

Neural style transfer has drawn considerable attention from both academic and industrial field. Although visual effect and efficiency have been significantly improved, existing methods are unable to coordinate spatial distribution of visual attention between the content image and stylized image, or…

Cited by 215PDFcodeScholar