← Search

Xiaoou Li

2 accepted papers

2025

BDC-CLIP: Brownian Distance Covariance for Adapting CLIP to Action Recognition

ICML 2025poster

Bridging contrastive language-image pre-training (CLIP) to video action recognition has attracted growing interest. Human actions are inherently rich in spatial and temporal contexts, involving dynamic interactions among people, objects, and the environment. Accurately recognizing actions requires e…

Cited by 0SourcePDFScholar
2025

ImagineFSL: Self-Supervised Pretraining Matters on Imagined Base Set for VLM-based Few-shot Learning

CVPR 2025highlight

Adapting CLIP models for few-shot recognition has recently attracted significant attention. Despite considerable progress, these adaptations remain hindered by the pervasive challenge of data scarcity. Text-to-image models, capable of generating abundant photorealistic labeled images, offer a promis…

Cited by 0SourcePDFScholar