Elevating Robust ASR By Decoupling Multi-Channel Speaker Separation and Speech Recognition
Despite the tremendous success of automatic speech recognition (ASR) with the introduction of deep learning, its performance is still unsatisfactory in many real-world multi-talker scenarios. Speaker separation excels in separating individual talkers but, as a frontend, it introduces processing arti…