2024
Audio-Journey: Open Domain Latent Diffusion Based Text-To-Audio Generation
ICASSP 2024accepted
Despite recent progress, machine learning (ML) models for open-domain audio generation need to catch up to generative models for image, text, speech, and music. The lack of massive open-domain audio datasets is the main reason for this performance gap; we overcome this challenge through a novel data…