2022
Learning Disentanglement with Decoupled Labels for Vision-Language Navigation
ECCV 2022poster
"Vision-and-Language Navigation (VLN) requires an agent to follow complex natural language instructions and perceive the visual environment for real-world navigation. Intuitively, we find that instruction disentanglement for each viewpoint along the agent’s path is critical for accurate navigation.…