ICASSP 2016accepted0 citations

Towards PLDA-RBM based speaker recognition in mobile environment: Designing stacked/deep PLDA-RBM systems

Andreas Nautsch, Hong Hao, Themos Stafylakis, Christian Rathgeb, Christoph Busch

Abstract

The vast majority of text-independent speaker recognition systems rely on intermediate-sized vectors (i-vectors), which are compared by probabilistic linear discriminant analysis (PLDA). This paper proposes a PLDA-alike approach with restricted Boltzmann machines for i-vector based speaker recognition: two deep architectures are presented and examined, which aim at suppressing channel effects and recovering speaker-discriminative information on back-ends trained on a small dataset. Experiments are carried out on the MOBIO SRE'13 database, which is a challenging and publicly available dataset for mobile speaker recognition with limited amounts of training data. The experiments show that the proposed system outperforms the baseline i-vector/PLDA approach by relative gains of 31% on female and 9% on male speakers in terms of half total error rate.

BibTeX
@inproceedings{icassp2016_towardspldarbmba,
  title = {Towards PLDA-RBM based speaker recognition in mobile environment: Designing stacked/deep PLDA-RBM systems},
  author = {Andreas Nautsch and Hong Hao and Themos Stafylakis and Christian Rathgeb and Christoph Busch},
  booktitle = {ICASSP 2016},
  year = {2016}
}