← Search

Zhongzhu Pu

1 accepted papers

2025

Praxis-VLM: Vision-Grounded Decision Making via Text-Driven Reinforcement Learning

NeurIPS 2025poster

Vision Language Models exhibit impressive performance for various tasks, yet they often lack the sophisticated situational reasoning required for complex decision-making. This paper shows that VLMs can achieve surprisingly strong decision-making performance when visual scenes are replaced by textual…

Cited by 0SourceScholar