← Search

Yuhao Kang

2 accepted papers

2026

SounDiT: Geo-Contextual Soundscape-to-Landscape Generation

CVPR 2026

Recent audio-to-image models have shown impressive performance in generating images of specific objects conditioned on their corresponding sounds. However, these models fail to reconstruct real-world landscapes conditioned on acoustic environments. To address this challenge, we present Geo-contextua

Cited by 0SourceScholar
2022

Modality Synergy Complement Learning with Cascaded Aggregation for Visible-Infrared Person Re-identification

ECCV 2022poster

"Visible-Infrared Re-Identification (VI-ReID) is challenging in image retrievals. The modality discrepancy will easily make huge intra-class variations. Most existing methods either bridge different modalities through modality-invariance or generate the intermediate modality for better performance.…