← Search

Tushar Choudhary

2 accepted papers

2024

LeGo-Drive: Language-enhanced Goal-oriented Closed-Loop End-to-End Autonomous Driving

IROS 2024poster

Existing Vision-Language Models (VLMs) produce long-term trajectory waypoints or directly control actions based on their perception input and language prompt. However, these VLMs are not explicitly aware of the constraints imposed by the scene or kinematics of the vehicle. As a result, the generated…

Cited by 3SourceScholar
2024

Talk2BEV: Language-enhanced Bird’s-eye View Maps for Autonomous Driving

ICRA 2024poster

This work introduces Talk2BEV, a large vision-language model (LVLM)1 interface for bird’s-eye view (BEV) maps commonly used in autonomous driving. While existing perception systems for autonomous driving scenarios have largely focused on a pre-defined (closed) set of object categories and driving sc…

Cited by 77SourcecodeScholar