2022
Textual Manifold-based Defense Against Natural Language Adversarial Examples
EMNLP 2022main
Despite the recent success of large pretrained language models in NLP, they are susceptible to adversarial examples. Concurrently, several studies on adversarial images have observed an intriguing property: the adversarial images tend to leave the low-dimensional natural data manifold. In this study…