← Search

Andy Tang

3 accepted papers

2026

Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control

RSS 2026poster

Pretrained vision-language models (VLMs) can make semantic and visual inferences across diverse settings, providing valuable common-sense priors for robotic control. However, effectively grounding this knowledge in robot behaviors remains an open challenge. Prior methods often employ a hierarchical …

Cited by 0SourceScholar
2025

Commonsense Reasoning for Legged Robot Adaptation with Vision-Language Models

ICRA 2025

Legged robots are physically capable of navigating a diverse variety of environments and overcoming a wide range of obstructions. For example, in a search and rescue mission, a legged robot could climb over debris, crawl through gaps, and navigate out of dead ends. However, the robot's controller ne

Cited by 21SourceScholar
2025

Learning Long-Context Diffusion Policies via Past-Token Prediction

CoRL 2025poster

Reasoning over long sequences of observations and actions is essential for many robotic tasks. Yet, learning effective long-context policies from demonstrations remains challenging. As context length increases, training becomes increasingly expensive due to rising memory demands, and policy perfor…

Cited by 0SourcecodeScholar