← Search

Nanfei Ye

2 accepted papers

2026

DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking

ICLR 2026poster

The advent of Vision-Language Models (VLMs) has significantly advanced end-to-end autonomous driving, demonstrating powerful reasoning abilities for high-level behavior planning tasks. However, existing methods are often constrained by a passive perception paradigm, relying solely on text-based reas…

Cited by 0SourcecodeScholar
2024

ModaLink: Unifying Modalities for Efficient Image-to-PointCloud Place Recognition

IROS 2024poster

Place recognition is an important task for robots and autonomous cars to localize themselves and close loops in pre-built maps. While single-modal sensor-based methods have shown satisfactory performance, cross-modal place recognition that retrieving images from a point-cloud database remains a chal…

Cited by 3SourcecodeScholar