← Search

Juan Zuluaga-Gomez

4 accepted papers

2025

Speech Data Selection for Efficient ASR Fine-Tuning using Domain Classifier and Pseudo-Label Filtering

ICASSP 2025accepted

In real-world speech data processing, the scarcity of annotated data and the abundance of unlabelled speech data present a significant challenge. To address this, we propose an efficient data selection pipeline for fine-tuning ASR models by generating pseudo-labels using WhisperX pipeline and select…

Cited by 6SourceScholar
2025

XLSR-Transducer: Streaming ASR for Self-Supervised Pretrained Models

ICASSP 2025accepted

Self-supervised pretrained models exhibit competitive performance in automatic speech recognition (ASR) on finetuning, even with limited in-domain supervised data. However, popular pretrained models are not suitable for streaming ASR because they are trained with full attention context. In this pape…

Cited by 0SourceScholar
2023

Effectiveness of Text, Acoustic, and Lattice-Based Representations in Spoken Language Understanding Tasks

ICASSP 2023accepted

In this paper, we perform an exhaustive evaluation of different representations to address the intent classification problem in a Spoken Language Understanding (SLU) setup. We benchmark three types of systems to perform the SLU intent detection task: 1) text-based, 2) lattice-based, and a novel 3) m…

Cited by 0SourceScholar
2022

A Two-Step Approach to Leverage Contextual Data: Speech Recognition in Air-Traffic Communications

ICASSP 2022accepted

Automatic Speech Recognition (ASR), as the assistance of speech communication between pilots and air-traffic controllers, can significantly reduce the complexity of the task and increase the reliability of transmitted information. ASR application can lead to a lower number of incidents caused by mis…

Cited by 0SourceScholar