MAID: A Conditional Diffusion Model for Long Music Audio Inpainting
Kaiyang Liu, Wendong Gan, Chenchen Yuan
Abstract
Recent works on long music audio inpainting have focused on unconditionally generating new segments to inpaint corrupted audio segments. However, the information about these segments may differ significantly from the original. To solve this problem, we propose MAID (Music Audio Inpainting DDPM), a model for music audio inpainting based on DDPM (Denoising Diffusion Probability Model). The model is capable of unconditional and conditional inpainting of music audio: (a) in the unconditional inpainting task, MAID is capable of inpainting gaps with a length between 200 ms and 1600 ms; (b) in the conditional inpainting task, the model can generate new segments similar to the original segments based on the piano-rolls corresponding to the gaps. Experiments show that MAID performs better than the baseline.
BibTeX
@inproceedings{icassp2023_maidaconditional,
title = {MAID: A Conditional Diffusion Model for Long Music Audio Inpainting},
author = {Kaiyang Liu and Wendong Gan and Chenchen Yuan},
booktitle = {ICASSP 2023},
year = {2023}
}