← Search

Kazushige Ouchi

3 accepted papers

2024

Audio Generation with Multiple Conditional Diffusion Model

AAAI 2024technical

Text-based audio generation models have limitations as they cannot encompass all the information in audio, leading to restricted controllability when relying solely on text. To address this issue, we propose a novel model that enhances the controllability of existing pre-trained text-to-audio models…

2024

Semi-Supervised Sound Event Detection with Local and Global Consistency Regularization

ICASSP 2024accepted

Learning meaningful frame-wise features on a partially labeled dataset is crucial to semi-supervised sound event detection. Prior works either maintain consistency on frame-level predictions or seek feature-level similarity among neighboring frames, which cannot exploit the potential of unlabeled da…

Cited by 0SourceScholar
2021

Syntactically Diverse Adversarial Network for Knowledge-Grounded Conversation Generation

EMNLP 2021finding

Generative conversation systems tend to produce meaningless and generic responses, which significantly reduce the user experience. In order to generate informative and diverse responses, recent studies proposed to fuse knowledge to improve informativeness and adopt latent variables to enhance the di…

Cited by 3SourcePDFScholar