2026
RoboOmni: Actions Are Just Another Modality for Your Vision-Language Models
ICML 2026poster
Integrating Vision-Language Models (VLMs) into robotics has facilitated the development of generalizable Vision-Language Action (VLA) policies. However, unified discrete frameworks lag behind decoupled continuous designs due to limitations in action chunking and temporal modeling. To address this, w…