SFCo-Nav: Efficient Zero-Shot Visual Language Navigation Via Collaboration of Slow LLM and Fast Attributed Graph Alignment
Recent advances in large vision-language models (VLMs) and large language models (LLMs) have enabled zero-shot approaches to Visual Language Navigation (VLN), where an agent follows natural language instructions using only ego perception and reasoning. However, existing zero‑shot methods typically c…