← Search

Shucong Zhang

5 accepted papers

2025

Linear Time Complexity Conformers with SummaryMixing for Streaming Speech Recognition

ICASSP 2025accepted

Automatic speech recognition (ASR) with an encoder equipped with self-attention, whether streaming or non-streaming, takes quadratic time in the length of the speech utterance. This slows down training and decoding, increase the cost, and limits the deployment of the ASR in constrained devices. Summ…

Cited by 0SourceScholar
2021

Train Your Classifier First: Cascade Neural Networks Training from Upper Layers to Lower Layers

ICASSP 2021accepted

Although the lower layers of a deep neural network learn features which are transferable across datasets, these layers are not transferable within the same dataset. That is, in general, freezing the trained feature extractor (the lower layers) and retraining the classifier (the upper layers) on the…

Cited by 0SourceScholar
2020

Learning Noise Invariant Features Through Transfer Learning For Robust End-to-End Speech Recognition

ICASSP 2020accepted

End-to-end models yield impressive speech recognition results on clean datasets while having inferior performance on noisy datasets. To address this, we propose transfer learning from a clean dataset (WSJ) to a noisy dataset (CHiME4) for connectionist temporal classification models. We argue that th…

Cited by 19SourceScholar