2026
Unifying Diffusion and Autoregression for Generalizable Vision-Language-Action Model
ICLR 2026poster
A central objective of manipulation policy design is to enable robots to comprehend human instructions and predict generalized actions in unstructured environments. Recent autoregressive vision-language-action (VLA) approaches discretize actions into bins to exploit the pretrained reasoning and gene…