← Search

Zhizhi Wang

1 accepted papers

2025

Data-centric NLP Backdoor Defense from the Lens of Memorization

NAACL 2025findings

Backdoor attack is a severe threat to the trustworthiness of DNN-based language models. In this paper, we first extend the definition of memorization of language models from sample-wise to more fine-grained sentence element-wise (e.g., word, phrase, structure, and style), and then point out that lan…

Cited by 3SourcePDFScholar