← Search

Zhenzhou Tan

1 accepted papers

2026

SkillNet: Hierarchical Skill Modeling for Compositional Generalization in Vision-Language Action Models

ICML 2026poster

Transfer across diverse task compositions and unseen behaviors remains a significant challenge for vision-language action (VLA) models. Skills are repeatable and atomic components for various tasks, and similarities shared with different skills provide evidence for transferability across behaviors. …

Cited by 0SourceScholar