Unsupervised and Semi-Supervised Few-Shot Acoustic Event Classification
Hsin-Ping Huang, Krishna C. Puvvada, Ming Sun, Chao Wang
Abstract
Few-shot Acoustic Event Classification (AEC) aims to learn a model to recognize novel acoustic events using very limited labeled data. Previous works utilize supervised pre-training as well as meta-learning approaches, which heavily rely on labeled data. Here, we study unsupervised and semi-supervised learning approaches for few-shot AEC. Our work builds upon recent advances in unsupervised representation learning introduced for speech recognition and language modeling. We learn audio representations from a large amount of unlabeled data, and use the resulting representations for few-shot AEC. We further extend our model in a semi-supervised fashion. Our unsupervised representation learning approach outperforms supervised pre-training methods, and our semi-supervised learning approach outperforms meta-learning methods for few-shot AEC. We also show that our work is more robust under domain mismatch.
BibTeX
@inproceedings{icassp2021_unsupervisedands,
title = {Unsupervised and Semi-Supervised Few-Shot Acoustic Event Classification},
author = {Hsin-Ping Huang and Krishna C. Puvvada and Ming Sun and Chao Wang},
booktitle = {ICASSP 2021},
year = {2021}
}