← Search

Srikanth Korse

5 accepted papers

2025

Parametric Object Coding in IVAS: Efficient Coding of Multiple Audio Objects at Low Bit Rates

ICASSP 2025accepted

The recently standardized 3GPP codec for Immersive Voice and Audio Services (IVAS) includes a parametric mode for efficiently coding multiple audio objects at low bit rates. In this mode, parametric side information is obtained from both the object metadata and the input audio objects. The side info…

Cited by 0SourceScholar
2022

A DNN Based Post-Filter to Enhance the Quality of Coded Speech in MDCT Domain

ICASSP 2022accepted

Frequency domain processing, and in particular the use of Modified Discrete Cosine Transform (MDCT), is the most widespread approach to audio coding. However, at low bitrates, audio quality, especially for speech, degrades drastically due to the lack of available bits to directly code the transform…

Cited by 0SourceScholar
2022

PostGAN: A GAN-Based Post-Processor to Enhance the Quality of Coded Speech

ICASSP 2022accepted

The quality of speech coded by transform coding is affected by various artefacts especially when bitrates to quantize the frequency components become too low. In order to mitigate these coding artefacts and enhance the quality of coded speech, a post-processor that relies on a-priori information tra…

Cited by 0SourceScholar
2018

GMM-Based Iterative Entropy Coding for Spectral Envelopes of Speech and Audio

ICASSP 2018accepted

Spectral envelope modelling is a central part of speech and audio codecs and is traditionally based on either vector quantization or scalar quantization followed by entropy coding. To bridge the coding performance of vector quantization with the low complexity of the scalar case, we propose an itera…

Cited by 0SourceScholar