← Search

Zicong Hu

2 accepted papers

2024

Boosting 3D Visual Grounding by Object-Centric Referring Network

IROS 2024poster

3D visual grounding is tasked with locating a specific object within a 3D scene, as described by a given textual reference. This task is challenging because it requires (1) the accurate recognition of various objects in a 3D scene and (2) the understanding of spatial relations in the description. Ho…

Cited by 0SourceScholar
2024

J-MAE: Jigsaw Meets Masked Autoencoders in X-Ray Security Inspection

ICASSP 2024accepted

The X-ray security inspection aims to identify any restricted items to protect public safety. Due to the lack of focus on unsupervised learning in this field, using pre-trained models on natural images leads to suboptimal results in downstream tasks. Previous works would lose the relative positional…

Cited by 0SourceScholar