2025
Praxis-VLM: Vision-Grounded Decision Making via Text-Driven Reinforcement Learning
NeurIPS 2025poster
Vision Language Models exhibit impressive performance for various tasks, yet they often lack the sophisticated situational reasoning required for complex decision-making. This paper shows that VLMs can achieve surprisingly strong decision-making performance when visual scenes are replaced by textual…