← Search

Chih-Wei Wu

5 accepted papers

2024

Looking Similar Sounding Different: Leveraging Counterfactual Cross-Modal Pairs for Audiovisual Representation Learning

CVPR 2024poster

Audiovisual representation learning typically relies on the correspondence between sight and sound. However there are often multiple audio tracks that can correspond with a visual scene. Consider for example different conversations on the same crowded street. The effect of such counterfactual pairs…

Cited by 4SourcePDFScholar
2024

Odaq: Open Dataset of Audio Quality

ICASSP 2024accepted

Research into the prediction and analysis of perceived audio quality is hampered by the scarcity of openly available datasets of audio signals accompanied by corresponding subjective quality scores. To address this problem, we present the Open Dataset of Audio Quality (ODAQ), a new dataset containin…

Cited by 0SourceScholar
2020

Orientation-aware Vehicle Re-identification with Semantics-guided Part Attention Network

ECCV 2020poster

Vehicle re-identification (re-ID) focuses on matching images of the same vehicle across different cameras. It is fundamentally challenging because differences between vehicles are sometimes subtle. While several studies incorporate spatial-attention mechanisms to help vehicle re-ID, they often requi…