← Search

Luca Della Libera

4 accepted papers

2026

FOCALCODEC-STREAM: STREAMING LOW-BITRATE SPEECH CODING VIA CAUSAL DISTILLATION

ICASSP 2026poster

Neural audio codecs are a fundamental component of modern generative audio pipelines. Although recent codecs achieve strong low-bitrate reconstruction and provide powerful representations for downstream tasks, most are non-streamable, limiting their use in real-time applications. We present FocalCod…

Cited by 0SourcePDFScholar
2025

FocalCodec: Low-Bitrate Speech Coding via Focal Modulation Networks

NeurIPS 2025poster

Large language models have revolutionized natural language processing through self-supervised pretraining on massive datasets. Inspired by this success, researchers have explored adapting these methods to speech by discretizing continuous audio into tokens using neural audio codecs. However, existin…

Cited by 0SourcecodeScholar
2024

Listenable Maps for Zero-Shot Audio Classifiers

NeurIPS 2024poster

Interpreting the decisions of deep learning models, including audio classifiers, is crucial for ensuring the transparency and trustworthiness of this technology. In this paper, we introduce LMAC-ZS (Listenable Maps for Zero-Shot Audio Classifiers), which, to the best of our knowledge, is the first d…

Cited by 4SourcePDFScholar
2024

Resource-Efficient Separation Transformer

ICASSP 2024accepted

Transformers have recently achieved state-of-the-art performance in speech separation. These models, however, are computationally demanding and require a lot of learnable parameters. This paper explores Transformer-based speech separation with a reduced computational cost. Our main contribution is t…

Cited by 0SourceScholar