2021
Turn the Combination Lock: Learnable Textual Backdoor Attacks via Word Substitution
ACL 2021long
Recent studies show that neural natural language processing (NLP) models are vulnerable to backdoor attacks. Injected with backdoors, models perform normally on benign examples but produce attacker-specified predictions when the backdoor is activated, presenting serious security threats to real-worl…