2025
Generating Vocals from Lyrics and Musical Accompaniment
ICASSP 2025accepted
In this work, we introduce AutoSing, a novel framework designed to generate diverse and high-quality singing voices from provided lyrics and musical accompaniment. AutoSing extends an existing semantic token-based text-to-speech approach by incorporating musical accompaniment as an additional condit…