2023
Impact of Adversarial Training on Robustness and Generalizability of Language Models
ACL 2023findings
Adversarial training is widely acknowledged as the most effective defense against adversarial attacks. However, it is also well established that achieving both robustness and generalization in adversarially trained models involves a trade-off. The goal of this work is to provide an in depth comparis…