← Search

Yuzhe Zhu

1 accepted papers

2026

MACS: Multi-source Audio-to-image Generation with Contextual Significance and Semantic Alignment

AAAI 2026technical

Propelled by the breakthrough in deep generative models, audio-to-image generation has emerged as a pivotal cross-modal task that converts complex auditory signals into rich visual representations. However, previous works only focus on single-source audio inputs for image generation, ignoring the mu

Cited by 0SourcePDFScholar