2025
Adversaries With Incentives: A Strategic Alternative to Adversarial Robustness
ICLR 2025poster
Adversarial training aims to defend against *adversaries*: malicious opponents whose sole aim is to harm predictive performance in any way possible. This presents a rather harsh perspective, which we assert results in unnecessarily conservative training. As an alternative, we propose to model oppone…