GOAT-Bench: A Benchmark for Multi-Modal Lifelong Navigation
The Embodied AI community has recently made significant strides in visual navigation tasks exploring targets from 3D coordinates objects language description and images. However these navigation models often handle only a single input modality as the target. With the progress achieved so far it is t…