← Search

Xuan Yao

4 accepted papers

2026

SC$^{2}$-WM: A Self-Correcting World Model with Closed-Loop Feedback for Vision-and-Language Navigation in Continuous Environments

ICML 2026poster

Vision-and-Language Navigation in Continuous Environments (VLN-CE) requires agents to make fine-grained navigation decisions under partial observability. However, most existing methods rely on open-loop execution, lacking mechanisms to detect and correct internal state drift during inference. We pro…

Cited by 0SourceScholar
2025

NavMorph: A Self-Evolving World Model for Vision-and-Language Navigation in Continuous Environments

ICCV 2025poster

Vision-and-Language Navigation in Continuous Environments (VLN-CE) requires agents to execute sequential navigation actions in complex environments guided by natural language instructions. Current approaches often struggle with generalizing to novel environments and adapting to ongoing changes durin…

2025

Unity in Diversity: Video Editing via Gradient-Latent Purification

CVPR 2025poster

Recently, text-driven video editing methods that optimize target latent representations have garnered significant attention and demonstrated promising results. However, these methods rely on self-supervised objectives to compute the gradients needed for updating latent representations, which inevita…

2024

Fast-Slow Test-Time Adaptation for Online Vision-and-Language Navigation

ICML 2024poster

The ability to accurately comprehend natural language instructions and navigate to the target location is essential for an embodied agent. Such agents are typically required to execute user instructions in an online manner, leading us to explore the use of unlabeled test samples for effective online…