← Search

Giorgio Becherini

6 accepted papers

2026

CLUTCH: Contextualized Language model for Unlocking Text-Conditioned Hand motion modelling in the wild

ICLR 2026poster

Hands play a central role in daily life, yet modeling natural hand motions remains underexplored. Existing methods that tackle text-to-hand-motion generation or hand animation captioning rely on studio-captured datasets with limited actions and contexts, making them costly to scale to “in-the-wild”…

Cited by 0SourceScholar
2026

MAMMA: Markerless Accurate Multi-person Motion Acquisition

CVPR 2026

We present MAMMA, a markerless motion-capture pipeline that accurately recovers SMPL-X parameters from multi-view video. Traditional motion-capture systems rely on physical markers. Although they offer high accuracy, their requirements of specialized hardware, manual marker placement, and extensive

Cited by 0SourcecodeScholar
2025

BEDLAM2.0: Synthetic humans and cameras in motion

NeurIPS 2025oral

Inferring 3D human motion from video remains a challenging problem with many applications. While traditional methods estimate the human in image coordinates, many applications require human motion to be estimated in world coordinates. This is particularly challenging when there is both human and cam…

Cited by 0SourceScholar
2025

Im2Haircut: Single-view Strand-based Hair Reconstruction for Human Avatars

ICCV 2025poster

We present a novel approach for 3D hair reconstruction from single photographs based on a global hair prior combined with local optimization. Capturing strand-based hair geometry from single photographs is challenging due to the variety and geometric complexity of hairstyles and the lack of ground t…

Cited by 0SourcePDFScholar
2024

EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling

CVPR 2024poster

We propose EMAGE a framework to generate full-body human gestures from audio and masked gestures encompassing facial local body hands and global movements. To achieve this we first introduce BEAT2 (BEAT-SMPLX-FLAME) a new mesh-level holistic co-speech dataset. BEAT2 combines a MoShed SMPL-X body wit…

2024

Emotional Speech-driven 3D Body Animation via Disentangled Latent Diffusion

CVPR 2024poster

Existing methods for synthesizing 3D human gestures from speech have shown promising results but they do not explicitly model the impact of emotions on the generated gestures. Instead these methods directly output animations from speech without control over the expressed emotion. To address this lim…