← Search

Zach Wood-Doughty

3 accepted papers

2024

Audio-Journey: Open Domain Latent Diffusion Based Text-To-Audio Generation

ICASSP 2024accepted

Despite recent progress, machine learning (ML) models for open-domain audio generation need to catch up to generative models for image, text, speech, and music. The lack of massive open-domain audio datasets is the main reason for this performance gap; we overcome this challenge through a novel data…

Cited by 0SourceScholar