← Search

Guanqi Zhan

7 accepted papers

2024

A General Protocol to Probe Large Vision Models for 3D Physical Understanding

NeurIPS 2024poster

Our objective in this paper is to probe large vision models to determine to what extent they ‘understand’ different physical properties of the 3D scene depicted in an image. To this end, we make the following contributions: (i) We introduce a general and lightweight protocol to evaluate whether feat…

2024

Broadcasting Support Relations Recursively from Local Dynamics for Object Retrieval in Clutters

RSS 2024poster

In our daily life, cluttered objects are everywhere, from scattered stationery and books cluttering the table to bowls and plates filling the kitchen sink. Retrieving a target object from clutters is an essential while challenging skill for robots, for the difficulty of safely manipulating an object…

Cited by 5SourcePDFScholar
2024

InstructNav: Zero-shot System for Generic Instruction Navigation in Unexplored Environment

CoRL 2024poster

Enabling robots to navigate following diverse language instructions in unexplored environments is an attractive goal for human-robot interaction. However, this goal is challenging because different navigation tasks require different strategies. The scarcity of instruction navigation data hinders tra…

Cited by 34SourceScholar
2023

Learning Environment-Aware Affordance for 3D Articulated Object Manipulation under Occlusions

NeurIPS 2023poster

Perceiving and manipulating 3D articulated objects in diverse environments is essential for home-assistant robots. Recent studies have shown that point-level affordance provides actionable priors for downstream manipulation tasks. However, existing works primarily focus on single-object scenarios wi…

Cited by 25SourcePDFScholar
2020

Generative 3D Part Assembly via Dynamic Graph Learning

NeurIPS 2020poster

Autonomous part assembly is a challenging yet crucial task in 3D computer vision and robotics. Analogous to buying an IKEA furniture, given a set of 3D parts that can assemble a single shape, an intelligent agent needs to perceive the 3D part geometry, reason to propose pose estimations for the inpu…

Cited by 100SourcePDFScholar