2026
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models
ICRA 2026poster
Vision-Language-Action (VLA) models such as OpenVLA, Octo, and π0 have shown strong generalization by leveraging large-scale demonstrations, yet their performance is still fundamentally constrained by the quality and coverage of supervised data. Reinforcement learning (RL) therefore provides a promi…