← Search

Yiqing Wang

3 accepted papers

2026

ArtHOI: Taming Foundation Models for Monocular 4D Reconstruction of Hand-Articulated-Object Interactions

CVPR 2026

Existing hand-object interactions (HOI) methods are largely limited to rigid objects, while 4D reconstruction methods of articulated objects generally require pre-scanning the object or even multi-view videos. It remains an unexplored but significant challenge to reconstruct 4D human-articulated-obj

Cited by 0SourcecodeScholar
2026

Boosting Medical Visual Understanding From Multi-Granular Language Learning

ICLR 2026poster

Recent advances in image-text pretraining have significantly enhanced visual understanding by aligning visual and textual representations. Contrastive Language-Image Pretraining (CLIP) has played a pivotal role in multimodal learning. However, its focus on single-label, single-granularity alignment…

Cited by 3SourcecodeScholar