← Search

Yatao Yang

1 accepted papers

2026

DUP: Detection-guided Unlearning for Backdoor Purification in Language Models

AAAI 2026technical

As backdoor attacks become more stealthy and robust, they reveal critical weaknesses in current defense strategies: detection methods often rely on coarse-grained feature statistics, and purification methods typically require full retraining or additional clean models. To address these challenges, w

Cited by 0SourcePDFScholar