← Search

Xiaoyan Li

8 accepted papers

2026

MARE: Multimodal Analogical Reasoning for Disease Evolution-Aware Radiology Report Generation

AAAI 2026technical

Radiology report generation from longitudinal medical data is critical for assessing disease progression and automating diagnostic workflows. While recent methods incorporate longitudinal information, they primarily rely on multimodal feature fusion, with limited capacity for explicit disease evolut

Cited by 0SourcePDFScholar
2026

TagaVLM: Topology-Aware Global Action Reasoning for Vision-Language Navigation

ICRA 2026poster

Vision-Language Navigation (VLN) presents a unique challenge for Large Vision-Language Models (VLMs) due to their inherent architectural mismatch: VLMs are primarily pretrained on static, disembodied vision-language tasks, which fundamentally clash with the dynamic, embodied, and spatially-structure…

2024

FastOcc: Accelerating 3D Occupancy Prediction by Fusing the 2D Bird’s-Eye View and Perspective View

ICRA 2024poster

In autonomous driving, 3D occupancy prediction outputs voxel-wise status and semantic labels for more comprehensive understandings of 3D scenes compared with traditional perception tasks, such as 3D object detection and bird’s-eye view (BEV) semantic segmentation. Recent researchers have extensively…

Cited by 35SourceScholar
2023

Center Focusing Network for Real-Time LiDAR Panoptic Segmentation

CVPR 2023poster

LiDAR panoptic segmentation facilitates an autonomous vehicle to comprehensively understand the surrounding objects and scenes and is required to run in real time. The recent proposal-free methods accelerate the algorithm, but their effectiveness and efficiency are still limited owing to the difficu…

2022

CPGNet: Cascade Point-Grid Fusion Network for Real-Time LiDAR Semantic Segmentation

ICRA 2022poster

LiDAR semantic segmentation essential for advanced autonomous driving is required to be accurate, fast, and easy-deployed on mobile platforms. Previous point-based or sparse voxel-based methods are far away from real-time applications since time-consuming neighbor searching or sparse 3D convolution…

Cited by 34SourcecodeScholar
2022

PRNet: Point-Range Fusion Network for Real-Time LiDAR Semantic Segmentation

IJCAI 2022poster

Accurate and real-time LiDAR semantic segmentation is necessary for advanced autonomous driving systems. To guarantee a fast inference speed, previous methods utilize the highly optimized 2D convolutions to extract features on the range view (RV), which is the most compact representation of the LiDA…

Cited by 2SourcePDFScholar
2022

Sequential Multi-View Fusion Network for Fast LiDAR Point Motion Estimation

ECCV 2022poster

"The LiDAR point motion estimation, including motion state prediction and velocity estimation, is crucial for understanding a dynamic scene in autonomous driving. Recent 2D projection-based methods run in real-time by applying the well-optimized 2D convolution networks on either the bird’s-eye view…

Cited by 3SourcePDFScholar