← Search

Jiawen Lin

2 accepted papers

2026

PC-CrossDiff: Point-Cluster Dual-Level Cross-Modal Differential Attention for Unified 3D Referring and Segmentation

AAAI 2026technical

3D Visual Grounding (3DVG) aims to localize the referent of natural language referring expressions through two core tasks: Referring Expression Comprehension (3DREC) and Segmentation (3DRES). While existing methods achieve high accuracy in simple, single-object scenes, they suffer from severe perfor

Cited by 0SourcePDFScholar
2026

UZ3DVG: Unaided Zero-Shot 3D Visual Grounding with Generated Language Conditions

CVPR 2026

Zero-Shot 3D Visual Grounding (Zero-Shot 3DVG) aims to localize target objects in 3D scenes from natural language descriptions without relying on instance-wise description annotations. Existing methods rely on extra 2D images during inference and/or require multi-turn interactions with large languag

Cited by 0SourcecodeScholar