2025
PAC Bench: Do Foundation Models Understand Prerequisites for Executing Manipulation Policies?
NeurIPS 2025poster
Vision-Language Models (VLMs) are increasingly pivotal for generalist robot manipulation, enabling tasks such as physical reasoning, policy generation, and failure detection. However, their proficiency in these high-level applications often assumes a deep understanding of low-level physical prerequi…