← Search

Qixin Zeng

1 accepted papers

2026

Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models

AAAI 2026technical

Vision-Language-Action (VLA) models based on flow matching have shown excellent performance in general-purpose robotic manipulation tasks. However, the action accuracy of these models on complex downstream tasks is unsatisfactory. One important reason is that these models rely solely on the post-tra

Cited by 0SourcePDFScholar