← Search

Zexi Zhang

1 accepted papers

2026

TagaVLM: Topology-Aware Global Action Reasoning for Vision-Language Navigation

ICRA 2026poster

Vision-Language Navigation (VLN) presents a unique challenge for Large Vision-Language Models (VLMs) due to their inherent architectural mismatch: VLMs are primarily pretrained on static, disembodied vision-language tasks, which fundamentally clash with the dynamic, embodied, and spatially-structure…