← Search

Xu Xiang

4 accepted papers

2023

Wespeaker: A Research and Production Oriented Speaker Embedding Learning Toolkit

ICASSP 2023accepted

Speaker modeling is essential for many related tasks, such as speaker recognition and speaker diarization. The dominant modeling approach is fixed-dimensional vector representation, i.e., speaker embedding. This paper introduces a research and production oriented speaker embedding learning toolkit,…

Cited by 0SourceScholar
2021

AISpeech-SJTU Accent Identification System for the Accented English Speech Recognition Challenge

ICASSP 2021accepted

This paper describes the AISpeech-SJTU system for the accent identification track of the Interspeech-2020 Accented English Speech Recognition Challenge. In this challenge track, only 160-hour accented English data collected from 8 countries and the auxiliary Librispeech dataset are provided for trai…

Cited by 0SourceScholar
2021

Unit Selection Synthesis Based Data Augmentation for Fixed Phrase Speaker Verification

ICASSP 2021accepted

Data augmentation is commonly used to help build a robust speaker verification system, especially in limited-resource case. However, conventional data augmentation methods usually focus on the diversity of acoustic environment, leaving the lexicon variation neglected. For text dependent speaker veri…

Cited by 0SourceScholar
2015

Recurrent neural network language model with structured word embeddings for speech recognition

ICASSP 2015accepted

Due to effective word context encoding and long-term context preserving, recurrent neural network language model (RNNLM) has attracted great interest by showing better performance over back-off n-gram models and feed-forward neural network language models (FNNLM). However, it still has the difficult…

Cited by 0SourceScholar