Rethinking Fusion: Disentangled Learning of Shared and Modality-Specific Information for Stance Detection
Multi-modal stance detection (MSD) aims to determine an author's stance toward a given target using both textual and visual content. While recent methods leverage multi-modal fusion and prompt-based learning, most fail to distinguish between modality-specific signals and cross-modal evidence, leadin…