2026
SounDiT: Geo-Contextual Soundscape-to-Landscape Generation
CVPR 2026
Recent audio-to-image models have shown impressive performance in generating images of specific objects conditioned on their corresponding sounds. However, these models fail to reconstruct real-world landscapes conditioned on acoustic environments. To address this challenge, we present Geo-contextua