2026
GAE: Unleashing Physical Potential of VLM with Generalizable Action Expert
ICML 2026poster
Vision-language models demonstrate strong reasoning and planning abilities, yet grounding these predictions into precise robot actions remains a central challenge. Existing Vision-Language-Action methods typically entangle reasoning and action generation, leading to limited generalization and costly…