2025
BDC-CLIP: Brownian Distance Covariance for Adapting CLIP to Action Recognition
ICML 2025poster
Bridging contrastive language-image pre-training (CLIP) to video action recognition has attracted growing interest. Human actions are inherently rich in spatial and temporal contexts, involving dynamic interactions among people, objects, and the environment. Accurately recognizing actions requires e…