← Search

Xiyuan Gao

5 accepted papers

2025

Asymmetric Reinforcing Against Multi-Modal Representation Bias

AAAI 2025technical

The strength of multimodal learning lies in its ability to integrate information from various sources, providing rich and comprehensive insights. However, in real-world scenarios, multi-modal systems often face the challenge of dynamic modality contributions, the dominance of different modalities ma…

2025

Intra-modal Relation and Emotional Incongruity Learning using Graph Attention Networks for Multimodal Sarcasm Detection

ICASSP 2025accepted

Sarcasm detection poses unique challenges due to the complex nature of sarcastic expressions often embedded across multiple modalities. Current methods frequently fall short in capturing the incongruent emotional cues that are essential for identifying sarcasm in multimodal contexts. In this paper,…

Cited by 0SourceScholar
2023

DecomFormer: Decompose Self-Attention Via Fourier Transform for VHR Aerial Image Scene Classification

ICASSP 2023accepted

Very high-resolution (VHR) aerial image scene classification is an essential task for aerial image understanding. Although transformer-based models have demonstrated strong ability in natural image classification, transformer-based methods on VHR aerial image tasks are still lack of concern because…

Cited by 0SourceScholar