ICASSP 2016accepted0 citations

Semi-autonomous data enrichment based on cross-task labelling of missing targets for holistic speech analysis

Yue Zhang, Yuxiang Zhou, Jie Shen, Björn W. Schuller

Abstract

In this work, we propose a novel approach for large-scale data enrichment, with the aim to address a major shortcoming of current research in computational paralinguistics, namely, looking at speaker attributes in isolation although strong interdependencies between them exist. The scarcity of multi-target databases, in which instances are labelled for different kinds of speaker characteristics, compounds this problem. The core idea of our work is to join existing data resources into one single holistic database with a multi-dimensional label space by using semi-supervised learning techniques to predict missing labels. In the proposed new Cross-Task Labelling (CTL) method, a model is first trained on the labelled training set of the selected databases for each individual task. Then, the trained classifiers are used for the crosslabelling of databases among each other. To exemplify the effectiveness of the `CTL' method, we evaluated it for likability, personality, and emotion recognition as representative tasks from the INTERSPEECH Computational Paralinguistics ChallengE (ComParE) series. The results show that `CTL' lays the foundation for holistic speech analysis by semi-autonomously annotating the existing databases, and expanding the multi-target label space at the same time, while achieving higher accuracy as the baseline performance of the challenges.

BibTeX
@inproceedings{icassp2016_semiautonomousda,
  title = {Semi-autonomous data enrichment based on cross-task labelling of missing targets for holistic speech analysis},
  author = {Yue Zhang and Yuxiang Zhou and Jie Shen and Björn W. Schuller},
  booktitle = {ICASSP 2016},
  year = {2016}
}