← Search

Vahid Noroozi

2 accepted papers

2024

Investigating End-to-End ASR Architectures for Long Form Audio Transcription

ICASSP 2024accepted

This paper presents an overview and evaluation of some of the end-to-end ASR models on long-form audio. We study three categories of Automatic Speech Recognition(ASR) models based on their core architecture: (1) convolutional, (2) convolutional with squeeze-and-excitation, and (3) convolutional mode…

Cited by 0SourceScholar
2024

Stateful Conformer with Cache-Based Inference for Streaming Automatic Speech Recognition

ICASSP 2024accepted

In this paper, we propose an efficient and accurate streaming speech recognition model based on the FastConformer architecture. We adapted the FastConformer architecture for streaming applications through: (1) constraining both the look-ahead and past contexts in the encoder, and (2) introducing an…

Cited by 0SourceScholar