2026
S²-VLA: State-Space Guided Vision-Language-Action Models for Long-Horizon Manipulation
IJCAI 2026
Vision-Language-Action (VLA) models have demonstrated strong capabilities in robotic manipulation, but their performance degrades significantly in long-horizon tasks due to cumulative error propagation. This limitation largely arises from static feature fusion mechanisms that rely on fixed weights t