← Search

Hema A. Murthy

8 accepted papers

2025

LIMMITS'25: Multilingual Streaming TTS With Neural Codecs for Indian Languages

ICASSP 2025accepted

This work provides a summary of the Multilingual streaming TTS with neural codecs for Indian languages challenge (LIMMITS’25), organized as part of the ICASSP 2025 signal processing grand challenge. Towards this, 278 hours of TTS data in 4 Indian languages - Gujarati, Indian English, Bhojpuri, and K…

Cited by 0SourceScholar
2023

Lightweight, Multi-Speaker, Multi-Lingual Indic Text-to-Speech

ICASSP 2023accepted

The Lightweight, Multi-speaker, Multi-lingual Indic Text-to-Speech (LIMMITS’23) challenge is organized as part of the ICASSP 2023 signal processing grand challenge. LIMMITS’23 aims at the development of a lightweight, multi-speaker, multi-lingual Text to Speech (TTS) model using datasets in Marathi,…

Cited by 0SourceScholar
2020

State-Based Transcription of Components of Carnatic Music

ICASSP 2020accepted

Automatic Carnatic Music (CM) transcription is an open problem in need of a standardized descriptive notation. The level of detail needed in a descriptive transcription makes it tedious to obtain ground truth by manual means. In this paper, we propose a novel state-based representation of the pitch…

Cited by 0SourceScholar
2019

An Empirical Study of Speech Processing in the Brain by Analyzing the Temporal Syllable Structure in Speech-input Induced EEG

ICASSP 2019accepted

Clinical applicability of electroencephalography (EEG) is well established, however the use of EEG as a choice for constructing brain computer interfaces to develop communication platforms is relatively recent. To provide more natural means of communication, there is an increasing focus on bringing…

Cited by 0SourceScholar
2019

Incremental Transfer Learning in Two-pass Information Bottleneck Based Speaker Diarization System for Meetings

ICASSP 2019accepted

The two-pass information bottleneck (TPIB) based speaker diarization system operates independently on different conversational recordings. TPIB system does not consider previously learned speaker discriminative information while di-arizing new conversations. Hence, the real time factor (RTF) of TPIB…

Cited by 0SourceScholar
2017

GDspike: An accurate spike estimation algorithm from noisy calcium fluorescence signals

ICASSP 2017accepted

Accurate estimation of spike train from calcium (Ca <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">2+</sup> ) fluorescence signals is challenging owing to significant fluctuations of fluorescence level. This paper proposes a non-model-based approach fo…

Cited by 10SourceScholar
2016

Significance of Pseudo-syllables in building better acoustic models for Indian English TTS

ICASSP 2016accepted

Signal processing based landmark detection is precise compared to HMM based alignment, primarily because the location of the landmark is not factored in the estimation of parameters. Acoustic cues for syllable boundaries are usually obtained by exploiting the inherent sonority characteristics of a s…

Cited by 12SourceScholar