2025
Making VLMs More Robot-Friendly: Self-Critical Distillation of Low-Level Procedural Reasoning
EMNLP 2025
Large language models (LLMs) have shown promise in robotic procedural planning, yet their human-centric reasoning often omits the low-level, grounded details needed for robotic execution. Vision-language models (VLMs) offer a path toward more perceptually grounded plans, but current methods either r