2024
CIF-T: A Novel CIF-Based Transducer Architecture for Automatic Speech Recognition
ICASSP 2024accepted
RNN-T models are widely used in ASR, which rely on the RNN-T loss to achieve length alignment between input audio and target sequence. However, the implementation complexity and the alignment-based optimization target of RNN-T loss lead to computational redundancy and a reduced role for predictor ne…