← Search

Haoyuan Yang

4 accepted papers

2025

BDC-CLIP: Brownian Distance Covariance for Adapting CLIP to Action Recognition

ICML 2025poster

Bridging contrastive language-image pre-training (CLIP) to video action recognition has attracted growing interest. Human actions are inherently rich in spatial and temporal contexts, involving dynamic interactions among people, objects, and the environment. Accurately recognizing actions requires e…

Cited by 0SourcePDFScholar
2025

ImagineFSL: Self-Supervised Pretraining Matters on Imagined Base Set for VLM-based Few-shot Learning

CVPR 2025highlight

Adapting CLIP models for few-shot recognition has recently attracted significant attention. Despite considerable progress, these adaptations remain hindered by the pervasive challenge of data scarcity. Text-to-image models, capable of generating abundant photorealistic labeled images, offer a promis…

Cited by 0SourcePDFScholar
2024

Self-Supervised Speaker Verification with Adaptive Threshold and Hierarchical Training

ICASSP 2024accepted

In self-supervised speaker verification, the quality of generated pseudo labels becomes a bottleneck for the performance. This work introduces a dynamic threshold within the iterative DIstillation with NO labels (DINO) framework. We employ a Gaussian Mixture Model (GMM) to model the loss distributio…

Cited by 0SourceScholar
2024

Wasserstein Distance Rivals Kullback-Leibler Divergence for Knowledge Distillation

NeurIPS 2024poster

Since pioneering work of Hinton et al., knowledge distillation based on Kullback-Leibler Divergence (KL-Div) has been predominant, and recently its variants have achieved compelling performance. However, KL-Div only compares probabilities of the corresponding category between the teacher and stud…

Cited by 1SourcePDFScholar