← Search

Ganesh Sivaraman

4 accepted papers

2025

Investigating voiced and unvoiced regions of speech for audio deepfake detection

ICASSP 2025accepted

Deep neural network based deepfake detection systems have achieved high levels of accuracy on benchmark datasets and competitions. However, most models lack interpretability. It is challenging to extract reasoning from the network that can convince the human evaluator to trust the decision. Humans o…

Cited by 11SourceScholar
2018

Smoothing Model Predictions Using Adversarial Training Procedures for Speech Based Emotion Recognition

ICASSP 2018accepted

Training discriminative classifiers involves learning a conditional distribution p(y <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">i</sup> |x <sub xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">i</sub> ), giv…

Cited by 0SourceScholar
2017

Joint modeling of articulatory and acoustic spaces for continuous speech recognition tasks

ICASSP 2017accepted

Articulatory information can effectively model variability in speech and can improve speech recognition performance under varying acoustic conditions. Learning speaker-independent articulatory models has always been challenging, as speaker-specific information in the articulatory and acoustic spaces…

Cited by 0SourceScholar