ICASSP 2023accepted0 citations

Pushing the Limits of Self-Supervised Speaker Verification using Regularized Distillation Framework

Yafeng Chen, Siqi Zheng, Hui Wang, Luyao Cheng, Qian Chen

Abstract

Training robust speaker verification systems without speaker labels has long been a challenging task. Previous studies observed a large performance gap between self-supervised and fully supervised methods. In this paper, we apply a non-contrastive self-supervised learning framework called DIstillation with NO labels (DINO) and propose two regularization terms applied to embeddings in DINO. One regularization term guarantees the diversity of the embeddings, while the other regularization term decorrelates the variables of each embedding. The effectiveness of various data augmentation techniques are explored, on both time and frequency domain. A range of experiments conducted on the VoxCeleb datasets demonstrate the superiority of the regularized DINO framework in speaker verification. Our method achieves the stateof-the-art speaker verification performance under a singlestage self-supervised setting on VoxCeleb.

BibTeX
@inproceedings{icassp2023_pushingthelimits,
  title = {Pushing the Limits of Self-Supervised Speaker Verification using Regularized Distillation Framework},
  author = {Yafeng Chen and Siqi Zheng and Hui Wang and Luyao Cheng and Qian Chen},
  booktitle = {ICASSP 2023},
  year = {2023}
}
Pushing the Limits of Self-Supervised Speaker Verification using Regularized Distillation Framework · ICASSP 2023