← Search

Xuyuan Han

1 accepted papers

2026

OpenDriveVLA: Towards End-to-end Autonomous Driving with Large Vision Language Action Model

AAAI 2026technical

We present OpenDriveVLA, a Vision-Language Action (VLA) model designed for end-to-end autonomous driving, built upon open-source large language models. OpenDriveVLA generates spatially-grounded driving actions by leveraging multimodal inputs, including both 2D and 3D instance-aware visual representa

Cited by 0SourcePDFScholar