← Search

Yasuhiro Minami

3 accepted papers

2024

CIF-RNNT: Streaming ASR Via Acoustic Word Embeddings with Continuous Integrate-and-Fire and RNN-Transducers

ICASSP 2024accepted

This paper introduces CIF-RNNT, a model that incorporates Continuous Integrate-and-Fire into RNN-Transducers (RNNTs) for streaming ASR via acoustic word embeddings (AWEs). CIF can dynamically compress long sequences into shorter ones, while RNNTs can produce multiple symbols given an input vector. W…

Cited by 2SourceScholar
2016

Speaker adaptive model based on Boltzmann machine for non-parallel training in voice conversion

ICASSP 2016accepted

In this paper, we present a voice conversion (VC) method that does not use any parallel data while training the model. VC is a technique where only speaker specific information in source speech is converted while keeping the phonological information unchanged. Most of the existing VC methods rely on…

Cited by 0SourceScholar