Multilingual Phonetic Dataset for Low Resource Speech Recognition
Xinjian Li, David R. Mortensen, Florian Metze, Alan W. Black
Abstract
Phone Recognition is one of the most important tasks in the field of multilingual speech recognition, especially for low-resource languages whose orthographies are not available. However, most speech recognition datasets so far only focus on high-resource languages, there are very few datasets available for low-resource languages, especially datasets with detailed phone annotation. In this work, we present a large multilingual phonetic dataset, which is preprocessed and aligned from the UCLA phonetic dataset. The dataset contains around 100 low-resource languages and 7000 utterances in total. This dataset would provide an ideal training/evaluation set for universal phone recognition.
BibTeX
@inproceedings{icassp2021_multilingualphon,
title = {Multilingual Phonetic Dataset for Low Resource Speech Recognition},
author = {Xinjian Li and David R. Mortensen and Florian Metze and Alan W. Black},
booktitle = {ICASSP 2021},
year = {2021}
}