← Search

I-hong Jhuo

2 accepted papers

2026

GOT-Edit: Geometry-Aware Generic Object Tracking via Online Model Editing

ICLR 2026poster

Human perception for effective object tracking in a 2D video stream arises from the implicit use of prior 3D knowledge combined with semantic reasoning. In contrast, most generic object tracking (GOT) methods primarily rely on 2D features of the target and its surroundings while neglecting 3D geomet…

Cited by 0SourcecodeScholar
2025

Generation and Comprehension Hand-in-Hand: Vision-guided Expression Diffusion for Boosting Referring Expression Generation and Comprehension

ICLR 2025poster

Referring expression generation (REG) and comprehension (REC) are vital and complementary in joint visual and textual reasoning. Existing REC datasets typically contain insufficient image-expression pairs for training, hindering the generalization of REC models to unseen referring expressions. More…

Cited by 0SourcePDFScholar