2026
ExPO-HM: Learning to Explain-then-Detect for Hateful Meme Detection
ICLR 2026poster
Hateful memes have emerged as a particularly challenging form of online abuse, motivating the development of automated detection systems. Most prior approaches rely on direct detection, producing only binary predictions. Such models fail to provide the context and explanations that real-world modera…