← Search

Chitralekha Gupta

7 accepted papers

2025

DroneAudioset: An Audio Dataset for Drone-based Search and Rescue

NeurIPS 2025poster

Unmanned Aerial Vehicles (UAVs) or drones, are increasingly used in search and rescue missions to detect human presence. Existing systems primarily leverage vision-based methods which are prone to fail under low-visibility or occlusion. Drone-based audio perception offers promise but suffers from ex…

Cited by 0SourceScholar
2025

MorphFader: Enabling Fine-grained Controllable Morphing with Text-to-Audio Models

ICASSP 2025accepted

Sound morphing is the process of gradually and smoothly transforming one sound into another to generate novel and perceptually hybrid sounds that simultaneously resemble both. Recently, diffusion-based text-to-audio models have produced high-quality sounds using text prompts. However, granularly con…

Cited by 0SourceScholar
2023

EMO-KNOW: A Large Scale Dataset on Emotion-Cause

EMNLP 2023short findings

Emotion-Cause analysis has attracted the attention of researchers in recent years. However, most existing datasets are limited in size and number of emotion categories. They often focus on extracting parts of the document that contain the emotion cause and fail to provide more abstractive, generaliz…

Cited by 0SourceScholar
2023

Towards Controllable Audio Texture Morphing

ICASSP 2023accepted

In this paper, we propose a data-driven approach to train a Generative Adversarial Network (GAN) conditioned on "soft-labels" distilled from the penultimate layer of an audio classifier trained on a target set of audio texture classes. We demonstrate that interpolation between such conditions or con…

Cited by 0SourceScholar
2022

Genre-Conditioned Acoustic Models for Automatic Lyrics Transcription of Polyphonic Music

ICASSP 2022accepted

Lyrics transcription of polyphonic music is challenging not only because the singing vocals are corrupted by the background music, but also because the background music and the singing style vary across music genres, such as pop, metal, and hip hop, which affects lyrics intelligibility of the song i…

Cited by 0SourceScholar
2020

Automatic Lyrics Alignment and Transcription in Polyphonic Music: Does Background Music Help?

ICASSP 2020accepted

Automatic lyrics alignment and transcription in polyphonic music are challenging tasks because the singing vocals are corrupted by the background music. In this work, we propose to learn music genre-specific characteristics to train polyphonic acoustic models. We first compare several automatic spee…

Cited by 0SourceScholar
2019

Automatic Lyrics-to-audio Alignment on Polyphonic Music Using Singing-adapted Acoustic Models

ICASSP 2019accepted

Lyrics-to-audio alignment is to automatically align the lyrical words with the mixed singing audio (singing voice+musical accompaniment). Such alignment can be achieved with an automatic speech recognition (ASR) system. We propose to adapt the acoustic model of a speech recognizer towards solo singi…

Cited by 0SourceScholar