← Search

Simon Rouard

4 accepted papers

2025

MusicGen-Stem: Multi-stem music generation and edition through autoregressive modeling

ICASSP 2025accepted

While most music generation models generate a mixture of stems (in mono or stereo), we propose to train a multi-stem generative model with 3 stems (bass, drums and other) that learn the musical dependencies between them. To do so, we train one specialized compression algorithm per stem to tokenize t…

Cited by 0SourceScholar
2024

An Independence-promoting Loss for Music Generation with Language Models

ICML 2024poster

Music generation schemes using language modeling rely on a vocabulary of audio tokens, generally provided as codes in a discrete latent space learnt by an auto-encoder. Multi-stage quantizers are often employed to produce these tokens, therefore the decoding strategy used for token prediction must b…

Cited by 3SourcePDFScholar