← Search

Zeyu Jiang

12 accepted papers

2026

CRAG: Can 3D Generative Models Help 3D Assembly?

ICML 2026poster

Most existing 3D assembly methods treat the problem as pure pose estimation, rearranging observed parts via rigid transformations. In contrast, human assembly naturally couples structural reasoning with holistic shape inference. Inspired by this intuition, we reformulate 3D assembly as a joint probl…

Cited by 0SourceScholar
2026

DexSinGrasp: Learning a Unified Policy for Dexterous Object Singulation and Grasping in Densely Cluttered Environments

RA-L 2026

Grasping objects in cluttered environments remains a fundamental yet challenging problem in robotic manipulation. While prior works have explored learning-based synergies between pushing and grasping for two-fingered grippers, few have leveraged the high degrees of freedom (DoF) in dexterous hands t

Cited by 3SourcecodeScholar
2026

DexSinGrasp: Learning a Unified Policy for Dexterous Object Singulation and Grasping in Densely Cluttered Environments

ICRA 2026poster

Grasping objects in cluttered environments remains a fundamental yet challenging problem in robotic manipulation. While prior works have explored learning-based synergies between pushing and grasping for two-fingered grippers, few have leveraged the high degrees of freedom (DoF) in dexterous hands t…

2026

Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes

CVPR 2026

Open-vocabulary 3D occupancy is vital for embodied agents, which need to understand complex indoor environments where semantic categories are abundant and evolve beyond fixed taxonomies. While recent work has explored open-vocabulary occupancy in outdoor driving scenarios, such methods transfer poor

Cited by 0SourcecodeScholar
2026

OrionEdit: Bridging Reference and Source Images for Generalized Cross-Image Editing

CVPR 2026

Multimodal image synthesis has made significant progress, yet most editing methods still rely on textual instructions, which are less direct than visual guidance. Recently, a new paradigm edits one image using another as reference, enabling more intuitive manipulation through visual exemplars. We fo

Cited by 0SourcecodeScholar
2026

SaferPath: Hierarchical Visual Navigation with Learned Guidance and Safety-Constrained Control

ICRA 2026poster

Visual navigation is a core capability for mobile robots, yet end-to-end learning-based methods often struggle with generalization and safety in unseen, cluttered, or narrow environments. These limitations are especially pronounced in dense indoor settings, where collisions are likely and end-to-end…

2025

Enhancing Scene Coordinate Regression With Efficient Keypoint Detection and Sequential Information

RA-L 2025

Scene Coordinate Regression (SCR) is a visual localization technique that utilizes deep neural networks (DNN) to directly regress 2D-3D correspondences for camera pose estimation. However, current SCR methods often face challenges in handling repetitive textures and meaningless areas due to their re

Cited by 3SourcecodeScholar
2025

GARF: Learning Generalizable 3D Reassembly for Real-World Fractures

ICCV 2025poster

3D reassembly is a challenging spatial intelligence task with broad applications across scientific domains. While large-scale synthetic datasets have fueled promising learning-based approaches, their generalizability to different domains is limited. Critically, it remains uncertain whether models tr…

Cited by 0SourcePDFScholar
2024

DIFFSC: Semantic Communication Framework With Enhanced Denoising Through Diffusion Probabilistic Models

ICASSP 2024accepted

In communication systems, the challenge of ensuring accurate data transmission across noisy channels remains paramount. While semantic communication shows potential in improving image transmission and reconstruction, existing methods still suffer from perceptual quality degradation in high-noise env…

Cited by 0SourceScholar
2018

Avoidance of High-Speed Obstacles Based on Velocity Obstacles

ICRA 2018poster

For obstacles moving with high speeds, existing motion planning methods can rarely guarantee collision avoidance. This paper proposes a viable two-period velocity obstacle algorithm where one period predicts potential collisions within a limited time horizon, and the second period foresees collision…

Cited by 20SourceScholar