← Search

Debjyoti Paul

3 accepted papers

2026

VowelPrompt: Hearing Speech Emotions from Text via Vowel-level Prosodic Augmentation

ICLR 2026poster

Emotion recognition in speech presents a complex multimodal challenge, requiring comprehension of both linguistic content and vocal expressivity, particularly prosodic features such as fundamental frequency, intensity, and temporal dynamics. Although large language models (LLMs) have shown promise i…

Cited by 0SourceScholar
2025

A Domain Adaptation Framework for Speech Recognition Systems with Only Synthetic data

ICASSP 2025accepted

We introduce DAS (Domain Adaptation with Synthetic data), a novel domain adaptation framework for pre-trained ASR model, designed to efficiently adapt to various language-defined domains without requiring any real data. In particular, DAS first prompts large language models (LLMs) to generate domain…

Cited by 0SourceScholar
2024

Recovering from Privacy-Preserving Masking with Large Language Models

ICASSP 2024accepted

Model adaptation is crucial to handle the discrepancy between proxy training data and actual users’ data received. To effectively perform adaptation, textual data of users is typically stored on servers or their local devices, where downstream natural language processing (NLP) models can be directly…

Cited by 0SourceScholar