← Search

Yexin Yang

4 accepted papers

2024

SeACo-Paraformer: A Non-Autoregressive ASR System with Flexible and Effective Hotword Customization Ability

ICASSP 2024accepted

Hotword customization is one of the concerned issues remained in ASR field - it is of value to enable users of ASR systems to customize names of entities, persons and other phrases to obtain better experience. The past few years have seen effective modeling strategies for ASR contextualization devel…

Cited by 0SourceScholar
2021

AISpeech-SJTU Accent Identification System for the Accented English Speech Recognition Challenge

ICASSP 2021accepted

This paper describes the AISpeech-SJTU system for the accent identification track of the Interspeech-2020 Accented English Speech Recognition Challenge. In this challenge track, only 160-hour accented English data collected from 8 countries and the auxiliary Librispeech dataset are provided for trai…

Cited by 0SourceScholar
2020

Text Adaptation for Speaker Verification with Speaker-Text Factorized Embeddings

ICASSP 2020accepted

Text mismatch between pre-collected data, either training data or enrollment data, and the actual test data can significantly hurt text-dependent speaker verification (SV) system performance. Although this problem can be solved by carefully collecting data with the target speech content, such data c…

Cited by 0SourceScholar
2019

Knowledge Distillation for Small Foot-print Deep Speaker Embedding

ICASSP 2019accepted

Deep speaker embedding learning is an effective method for speaker identity modelling. Very deep models such as ResNet can achieve remarkable results but are usually too computationally expensive for real applications with limited resources. On the other hand, simply reducing model size is likely to…

Cited by 0SourceScholar