← Search

Debmalya Chakrabarty

4 accepted papers

2025

AMuSE: Attentive Multilingual Speech Encoding for Zero-Prior ASR

ICASSP 2025accepted

Multilingual ASR offers training, deployment and overall performance benefits, but models trained via simple data pooling are known to suffer from cross-lingual interference. Oracle language information (exact-prior) and language-specific parameters are usually leveraged to overcome this, but such a…

Cited by 0SourceScholar
2022

Multi-Modal Pre-Training for Automated Speech Recognition

ICASSP 2022accepted

Traditionally, research in automated speech recognition has focused on local-first encoding of audio representations to predict the spoken phonemes in an utterance. Unfortunately, approaches relying on such hyper-local information tend to be vulnerable to both local-level corruption (such as audio-f…

Cited by 0SourceScholar
2019

Joint Acoustic and Class Inference for Weakly Supervised Sound Event Detection

ICASSP 2019accepted

Sound event detection is a challenging task, especially for scenes with multiple simultaneous events. While event classification methods tend to be fairly accurate, event localization presents additional challenges, especially when large amounts of labeled data are not available. Task4 of the 2018 D…

Cited by 0SourceScholar