← Search

Zhenhua Tang

4 accepted papers

2026

Glimpse: Geometry Learning of Multi-scale Structural Priors for 3D Pose Estimation

ICML 2026poster

Monocular 3D human pose estimation is fundamentally challenged by severe occlusion and inherent depth ambiguity. To address this, we propose Glimpse, a framework that learns robust 3D poses by explicitly modeling anatomical geometry from a single image. We recast the problem as geometry learning of …

Cited by 0SourceScholar
2026

SNS-Grasp: Semantic-guided Noise Scaling for Grasp Generation

AAAI 2026technical

While diffusion models show promise for intent-based grasp generation, their isotropic noise schedules struggle with joint-specific sensitivity and task-aware variability. This limitation leads to grasps with suboptimal semantic alignment or physical feasibility. To address this challenge, we propos

Cited by 0SourcePDFScholar
2025

RAGG: Retrieval-Augmented Grasp Generation Model

AAAI 2025technical

Intent-based grasp generation inherently involves challenges such as manipulation ambiguity and modality gaps. To address these, we propose a novel Retrieval-Augmented Grasp Generation model (RAGG). Our key insight is that when humans manipulate new objects, they initially mimic the interaction patt…

Cited by 0SourcePDFScholar