LatentVLA: Taming Latent Space for Generalizable and Long-Horizon Bimanual Manipulation
Current paradigms for robotic imitation learning face a stark trade-off between the motion fidelity of diffusion models and the data scalability of inverse dynamics models. The latter, while scalable, often learns a latent action space disconnected from physical reality. This flaw leads to critical