2025
ChunkFormer: Masked Chunking Conformer For Long-Form Speech Transcription
ICASSP 2025accepted
Deploying ASR models at an industrial scale poses significant challenges in hardware resource management, especially for long-form transcription tasks where audio may last for hours. Large Conformer models, despite their capabilities, are limited to processing only 15 minutes of audio on an 80GB GPU…