2025
Language Decoupling with Fine-grained Knowledge Guidance for Referring Multi-object Tracking
ICCV 2025poster
Referring Multi-Object Tracking (RMOT) aims to detect and track specific objects based on natural language expressions. Previous methods typically rely on sentence-level vision-language alignment, often failing to exploit fine-grained linguistic cues that are crucial for distinguishing objects with…