2026
Dual-View Predictive Diffusion: Lightweight Speech Enhancement via Spectrogram-Image Synergy
ICML 2026poster
Diffusion models have recently set new benchmarks in Speech Enhancement (SE). However, most existing score-based models treat speech spectrograms merely as generic 2D images, applying uniform processing that ignores the intrinsic structural sparsity of audio, which results in inefficient spectral re…