2026
LARA: Latent Action Representation Alignment for Vision-Language-Action Models
ICML 2026poster
Visual-language action (VLA) models enable robots to predict actions directly from observations and language instructions, but their performance depends on large-scale, high-quality data and is limited by the scarcity of real-world robot action datasets. To facilitate VLA model learning with abundan…