← Search

Yujia Shi

3 accepted papers

2026

Beyond Skeletons: Learning Animation Directly from Driving Videos with Same2X Training Strategy

ICLR 2026poster

Human image animation aims to generate a video from a static reference image, guided by pose information extracted from a driving video. Existing approaches often rely on pose estimators to extract intermediate representations, but such signals are prone to errors under occlusion or complex poses. B…

Cited by 0SourcecodeScholar
2026

EgoTactile: Learning Grasp Pressure for Everyday Objects from Egocentric Video

ICML 2026spotlight

Estimating full-hand grasp pressure from egocentric video is critical for immersive VR and robotic manipulation, yet dense tactile sensing often relies on intrusive hardware. Existing vision-based methods predominantly rely on planar surfaces or fingertip contacts, failing to generalize to complex 3…

Cited by 0SourcecodeScholar
2020

Not only Look, but also Listen: Learning Multimodal Violence Detection under Weak Supervision

ECCV 2020poster

but also Listen: Learning Multimodal Violence Detection under Weak Supervision","Violence detection has been studied in computer vision for years. However, previous work are either superficial, e.g., classification of short-clips, and the single scenario, or undersupplied, e.g., the single modality,…