← Search

Karan Nathwani

3 accepted papers

2025

Exploiting Wavelet Scattering Transform & Squeeze-Excitation Blocks with Cross-Modal Attention for Multi-modal Emotion Recognition

ICASSP 2025accepted

Multi-modal emotion recognition (MER) is crucial for improving human-computer interaction. Convolutional neural networks (CNNs) are the mainstream for MER tasks, but they require large databases, extensive memory, and significant energy, limiting their practical use. This paper proposes a novel MER…

Cited by 0SourceScholar
2023

Exploiting Sparse Recovery Algorithms for Semi-Supervised Training of Deep Neural Networks for Direction-of-Arrival Estimation

ICASSP 2023accepted

This paper proposes a semi-supervised training approach for a direction-of-arrival (DoA) estimation based on a convolutional neural network (CNN). We apply a sparse recovery algorithm called optMGD-ℓ <inf xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</in…

Cited by 0SourceScholar
2016

Formant shifting for speech intelligibility improvement in car noise environment

ICASSP 2016accepted

In this paper, we propose a novel approach aiming at improving the intelligibility of speech in the context of in-car applications. Speech produced in noisy environments is subject to the Lombard effect which gathers a number of voice transformation effects compared to the speech produced in calm en…

Cited by 0SourceScholar