← Search

Yuyue Wang

3 accepted papers

2025

VAFlow: Video-to-Audio Generation with Cross-Modality Flow Matching

ICCV 2025poster

Video-to-audio (V2A) generation aims to synthesize temporally aligned, realistic sounds for silent videos, a critical capability for immersive multimedia applications. Current V2A methods, predominantly based on diffusion or flow models, rely on suboptimal noise-to-audio paradigms that entangle cros…