← Search

Jianhui Zhang

6 accepted papers

2026

AR-Nav Benchmark: Augmented Reality Navigation with Vision and Language

AAAI 2026technical

Augmented Reality (AR) navigation has emerged as a transformative tool for spatial intelligence, enabling users to interactively explore complex environments through wearable and mobile AR devices. However, current AR navigation systems struggle with low indoor localization accuracy, weak semantic u

Cited by 0SourcePDFScholar
2026

Chain-of-Search: Parameter-Efficient Reasoning for Zero-Shot Object Navigation

AAAI 2026technical

Zero-shot object navigation tasks agents with locating target objects in unseen environments—a core capability of embodied intelligence. While recent vision-language navigation methods leverage Large Language Models (LLMs) for multimodal reasoning, they suffer from two key limitations: (1) semantic

Cited by 0SourcePDFScholar
2025

CamPoint: Boosting Point Cloud Segmentation with Virtual Camera

CVPR 2025poster

Local features aggregation and global information perception are the fundamental to point cloud segmentation. However, existing works often fall short in effectively identifying semantic relevant neighbors and face challenges in endowing each point with high-level information. Here, we propose CamPo…

Cited by 0SourcePDFScholar
2025

Ultra High-Resolution Image Inpainting with Patch-Based Content Consistency Adapter

ICCV 2025poster

In this work, we present Patch-Adapter, an effective framework for high-resolution text-guided image inpainting. Unlike existing methods limited to lower resolutions, our approach achieves 4K+ resolution while maintaining precise content consistency and prompt alignment--two critical challenges in i…

2024

Sparse Multi-Relational Graph Convolutional Network for Multi-type Object Trajectory Prediction

IJCAI 2024poster

Object trajectory prediction is a hot research issue with wide applications in video surveillance and autonomous driving. The previous studies consider the interaction sparsity mainly among the pedestrians instead of multi-type of objects, which brings new types of interactions and consequently supe…

Cited by 1SourcePDFScholar
2024

UAV First-Person Viewers Are Radiance Field Learners

ECCV 2024poster

"First-Person-View (FPV) holds immense potential for revolutionizing the trajectory of Unmanned Aerial Vehicles (UAVs), offering an exhilarating avenue for navigating complex building structures. Yet, traditional Neural Radiance Field (NeRF) methods face challenges such as sampling single points per…

Cited by 0SourcePDFScholar