2025
QueryAdapter: Rapid Adaptation of Vision-Language Models in Response to Natural Language Queries
IROS 2025
A domain shift exists between the large-scale, internet data used to train a Vision-Language Model (VLM) and the raw image streams collected by a robot. Existing adaptation strategies require the definition of a closed-set of classes, which is impractical for a robot that must respond to diverse nat