2024
GeoReasoner: Geo-localization with Reasoning in Street Views using a Large Vision-Language Model
ICML 2024poster
This work tackles the problem of geo-localization with a new paradigm using a large vision-language model (LVLM) augmented with human inference knowledge. A primary challenge here is the scarcity of data for training the LVLM - existing street-view datasets often contain numerous low-quality images…