2026
LAOF: Robust Latent Action Learning with Optical Flow Constraints
CVPR 2026
Learning latent actions from large-scale videos is crucial for the pre-training of scalable embodied foundation models, yet existing methods often struggle with action-irrelevant distractors. Although incorporating action supervision can alleviate these distractions, its effectiveness is restricted