← Search

Swarup Ranjan Behera

2 accepted papers

2025

Investigating Prosodic Signatures via Speech Pre-Trained Models for Audio Deepfake Source Attribution

ACL 2025finding

In this work, we investigate various state-of-the-art (SOTA) speech pre-trained models (PTMs) for their capability to capture prosodic sig-natures of the generative sources for audio deepfake source attribution (ADSD). These prosodic characteristics can be considered oneof major signatures for ADSD,…

Cited by 0SourcePDFScholar
2025

Strong Alone, Stronger Together: Synergizing Modality-Binding Foundation Models with Optimal Transport for Non-Verbal Emotion Recognition

ICASSP 2025accepted

In this study, we investigate multimodal foundation models (MFMs) for emotion recognition from non-verbal sounds. We hypothesize that MFMs, with their joint pre-training across multiple modalities, will be more effective in non-verbal sounds emotion recognition (NVER) by better interpreting and diff…

Cited by 0SourceScholar