2025
T2ICount: Enhancing Cross-modal Understanding for Zero-Shot Counting
CVPR 2025highlight
Zero-shot object counting aims to count instances of arbitrary object categories specified by text descriptions. Existing methods typically rely on vision-language models like CLIP, but often exhibit limited sensitivity to text prompts. We present T2ICount, a diffusion-based framework that leverages…