← Search

Lirong Yang

4 accepted papers

2024

Eliminating Cross-modal Conflicts in BEV Space for LiDAR-Camera 3D Object Detection

ICRA 2024poster

Recent 3D object detectors typically utilize multi-sensor data and unify multi-modal features in the shared bird’s-eye view (BEV) representation space. However, our empirical findings indicate that previous methods have limitations in generating fusion BEV features free from cross-modal conflicts. T…

Cited by 12SourcecodeScholar
2024

Global-Local Collaborative Inference with LLM for Lidar-Based Open-Vocabulary Detection

ECCV 2024poster

"Open-Vocabulary Detection (OVD) is the task of detecting all interesting objects in a given scene without predefined object classes. Extensive work has been done to deal with the OVD for 2D RGB images, but the exploration of 3D OVD is still limited. Intuitively, lidar point clouds provide 3D inform…

2023

Adaptive Zone-Aware Hierarchical Planner for Vision-Language Navigation

CVPR 2023poster

The task of Vision-Language Navigation (VLN) is for an embodied agent to reach the global goal according to the instruction. Essentially, during navigation, a series of sub-goals need to be adaptively set and achieved, which is naturally a hierarchical navigation process. However, previous methods l…

2020

CenterMask: Single Shot Instance Segmentation With Point Representation

CVPR 2020poster

In this paper, we propose a single-shot instance segmentation method, which is simple, fast and accurate. There are two main challenges for one-stage instance segmentation: object instances differentiation and pixel-wise feature alignment. Accordingly, we decompose the instance segmentation into two…

Cited by 108PDFScholar