← Search

Aswin Shanmugam Subramanian

5 accepted papers

2023

Reverberation as Supervision For Speech Separation

ICASSP 2023accepted

This paper proposes reverberation as supervision (RAS), a novel unsupervised loss function for single-channel reverberant speech separation. Prior methods for unsupervised separation required the synthesis of mixtures of mixtures or assumed the existence of a teacher model, making them difficult to…

Cited by 0SourceScholar
2021

Directional ASR: A New Paradigm for E2E Multi-Speaker Speech Recognition with Source Localization

ICASSP 2021accepted

This paper proposes a new paradigm for handling far-field multi-speaker data in an end-to-end (E2E) neural network manner, called directional automatic speech recognition (D-ASR), which explicitly models source speaker locations. In D-ASR, the azimuth angle of the sources with respect to the microph…

Cited by 0SourceScholar
2020

Attention-Based ASR with Lightweight and Dynamic Convolutions

ICASSP 2020accepted

End-to-end (E2E) automatic speech recognition (ASR) with sequence-to-sequence models has gained attention because of its simple model training compared with conventional hidden Markov model based ASR. Recently, several studies report the state-of-the-art E2E ASR results obtained by Transformer. Comp…

Cited by 0SourceScholar
2020

Far-Field Location Guided Target Speech Extraction Using End-to-End Speech Recognition Objectives

ICASSP 2020accepted

Target speech extraction is a specific case of source separation where an auxiliary information like the location or some pre-saved anchor speech examples of the target speaker is used to resolve the permutation ambiguity. Traditionally such systems are optimized based on signal reconstruction objec…

Cited by 0SourceScholar