← Search

Karan Wanchoo

1 accepted papers

2022

Cross-Modal Map Learning for Vision and Language Navigation

CVPR 2022poster

We consider the problem of Vision-and-Language Navigation (VLN). The majority of current methods for VLN are trained end-to-end using either unstructured memory such as LSTM, or using cross-modal attention over the egocentric observations of the agent. In contrast to other works, our key insight is…

Cited by 83PDFcodeScholar