2025
Beyond WER: Probing Whisper’s Sub‐token Decoder Across Diverse Language Resource Levels
EMNLP 2025
While large multilingual automatic speech recognition (ASR) models achieve remarkable performance, the internal mechanisms of the end-to-end pipeline, particularly concerning fairness and efficacy across languages, remain underexplored. This paper introduces a fine-grained analysis of Whisper’s mult