Combining Conformer and Dual-Path-Transformer Networks for Single Channel Noisy Reverberant Speech Separation
Separation of overlapping speakers remains an active area of speech technology research. Many deep neural network (DNN) separation models propose modelling local and global temporal context separately using alternating DNN layers. Two such models are SepFormer and TD-Conformer. The largest configura…