← Search

Bunlong Lay

3 accepted papers

2024

EMOCONV-Diff: Diffusion-Based Speech Emotion Conversion for Non-Parallel and in-the-Wild Data

ICASSP 2024accepted

Speech emotion conversion is the task of converting the expressed emotion of a spoken utterance to a target emotion while preserving the lexical content and speaker identity. While most existing works in speech emotion conversion rely on acted-out datasets and parallel data samples, in this work we…

Cited by 0SourceScholar
2024

Single and Few-Step Diffusion for Generative Speech Enhancement

ICASSP 2024accepted

Diffusion models have shown promising results in single-channel speech enhancement, using a task-adapted diffusion process for the conditional generation of clean speech given a noisy mixture. However, at test time, the neural network used for score estimation is called multiple times to solve the i…

Cited by 0SourceScholar
2023

Speech Signal Improvement Using Causal Generative Diffusion Models

ICASSP 2023accepted

In this paper, we present a causal speech signal improvement system that is designed to handle different types of distortions. The method is based on a generative diffusion model which has been shown to work well in scenarios with missing data and non-linear corruptions. To guarantee causal processi…

Cited by 0SourceScholar