Joint Training of Complex Ratio Mask Based Beamformer and Acoustic Model for Noise Robust Asr
In this paper, we present a joint training framework between the multi-channel beamformer and the acoustic model for noise robust automatic speech recognition (ASR). The complex ratio mask (CRM), demonstrated to be more effective than the ideal ratio mask (IRM), is proposed to estimate the covarianc…