ViTPrompt: Training-Free Prompt Refinement with Visual Tokens for Open-Vocabulary Detection
Test-Time Adaptive Object Detection (TTAOD) aims to maintain detection performance under distribution shifts without retraining. While recent vision-language models enable open-vocabulary detection, existing TTAOD methods--whether closed-set or open-vocabulary--focus exclusively on improving classif