GenCo: A Dual VLM Generate-Correct Framework for Adaptive Peg-in-Hole Robotics
Recent advances in Vision Language Models (VLMs) have enhanced their application in robotics, encompassing both high-level task planning and low-level action control. Despite their strong performance across various robotic tasks, even for zero-shot scenarios, most VLM applications remain open-loop,