← Search

Anran Lin

2 accepted papers

2024

Interactive Navigation in Environments with Traversable Obstacles Using Large Language and Vision-Language Models

ICRA 2024poster

This paper proposes an interactive navigation framework by using large language and vision-language models, allowing robots to navigate in environments with traversable obstacles. We utilize the large language model (GPT-3.5) and the open-set Vision-language Model (Grounding DINO) to create an actio…

Cited by 12SourceScholar
2021

Refer-It-in-RGBD: A Bottom-Up Approach for 3D Visual Grounding in RGBD Images

CVPR 2021poster

Grounding referring expressions in RGBD image has been an emerging field. We present a novel task of 3D visual grounding in single-view RGBD image where the referred objects are often only partially scanned due to occlusion. In contrast to previous works that directly generate object proposals for g…

Cited by 43PDFScholar