← Search

Ahmed Helmy

4 accepted papers

2026

LiveGesture: Streamable Co-Speech Gesture Generation Model

CVPR 2026

We propose LiveGesture, the first fully streamable, speech-driven full-body gesture generation framework that operates with zero look-ahead and supports arbitrary sequence length. Unlike existing co-speech gesture methods--which are designed for offline generation and either treat body regions indep

Cited by 0SourceScholar
2026

MS-Temba: Multi-Scale Temporal Mamba for Understanding Long Untrimmed Videos

CVPR 2026

Temporal Action Detection (TAD) in untrimmed videos poses significant challenges, particularly for Activities of Daily Living (ADL) requiring models to (1) process long-duration videos, (2) capture temporal variations in actions, and (3) simultaneously detect dense overlapping actions. Existing CNN

Cited by 0SourcecodeScholar
2026

Walk Before You Dance: High-fidelity and Editable Dance Synthesis via Generative Masked Motion Prior

AAAI 2026technical

Recent advances in dance generation have enabled the automatic synthesis of 3D dance motions. However, existing methods still face significant challenges in simultaneously achieving high realism, precise dance-music synchronization, diverse motion expression, and physical plausibility. To address th

Cited by 0SourcePDFScholar
2025

MaskHand: Generative Masked Modeling for Robust Hand Mesh Reconstruction in the Wild

ICCV 2025poster

Reconstructing a 3D hand mesh from a single RGB image is challenging due to complex articulations, self-occlusions, and depth ambiguities. Traditional discriminative methods, which learn a deterministic mapping from a 2D image to a single 3D mesh, often struggle with the inherent ambiguities in 2D-t…