2026
Locate then Correct: Debiasing Attention Heads in CLIP
ICML 2026poster
Multimodal models like CLIP have gained significant attention due to their remarkable zero-shot performance across various tasks. However, studies have revealed that CLIP can inadvertently learn spurious associations between target variables and confounding factors. To address this, we introduce \te…