2025
Sketch2Sound: Controllable Audio Generation via Time-Varying Signals and Sonic Imitations
ICASSP 2025accepted
We present Sketch2Sound, a generative audio model capable of creating high-quality sounds from a set of interpretable time-varying control signals: loudness, brightness, and pitch, as well as text prompts. Sketch2Sound can synthesize arbitrary sounds from sonic imitations (i.e., a vocal imitation or…