← Search

Travis M. Bartley

3 accepted papers

2025

HAINAN: Fast and Accurate Transducer for Hybrid-Autoregressive ASR

ICLR 2025poster

We present Hybrid-Autoregressive INference TrANsducers (HAINAN), a novel architecture for speech recognition that extends the Token-and-Duration Transducer (TDT) model. Trained with randomly masked predictor network outputs, HAINAN supports both autoregressive inference with all network components a…

Cited by 0SourcePDFScholar
2024

A Chat about Boring Problems: Studying GPT-Based Text Normalization

ICASSP 2024accepted

Text normalization - the conversion of text from written to spoken form - is traditionally assumed to be an ill-formed task for language modeling. In this work, we argue otherwise. We empirically show the capacity of Large-Language Models (LLM) for text normalization in few-shot scenarios. Combining…

Cited by 0SourceScholar
2023

Accidental Learners: Spoken Language Identification in Multilingual Self-Supervised Models

ICASSP 2023accepted

In this paper, we extend previous self-supervised approaches for language identification by experimenting with Conformer based architecture in a multilingual pre-training paradigm. We find that pre-trained speech models optimally encode language discriminatory information in lower layers. Further, w…

Cited by 0SourceScholar