Self-Information Guided Speech Segmentation for Efficient Streaming ASR
Unlike modern streaming Automatic Speech Recognition (ASR) systems which segment speech into fixed-length chunks for decoding, humans perceive speech in variable-length units of information. This paper proposes a novel method that leverages self-information, a measure of the information contained wi…