NeurIPS 2019poster292 citations

Defense Against Adversarial Attacks Using Feature Scattering-based Adversarial Training

Haichao Zhang, Jianyu Wang

Abstract

We introduce a feature scattering-based adversarial training approach for improving model robustness against adversarial attacks. Conventional adversarial training approaches leverage a supervised scheme (either targeted or non-targeted) in generating attacks for training, which typically suffer from issues such as label leaking as noted in recent works. Differently, the proposed approach generates adversarial images for training through feature scattering in the latent space, which is unsupervised in nature and avoids label leaking. More importantly, this new approach generates perturbed images in a collaborative fashion, taking the inter-sample relationships into consideration. We conduct analysis on model robustness and demonstrate the effectiveness of the proposed approach through extensively experiments on different datasets compared with state-of-the-art approaches.

BibTeX
@inproceedings{NEURIPS2019_d8700cbd,
 author = {Zhang, Haichao and Wang, Jianyu},
 booktitle = {Advances in Neural Information Processing Systems},
 editor = {H. Wallach and H. Larochelle and A. Beygelzimer and F. d\textquotesingle Alch\'{e}-Buc and E. Fox and R. Garnett},
 pages = {},
 publisher = {Curran Associates, Inc.},
 title = {Defense Against Adversarial Attacks Using Feature Scattering-based Adversarial Training},
 url = {https://proceedings.neurips.cc/paper_files/paper/2019/file/d8700cbd38cc9f30cecb34f0c195b137-Paper.pdf},
 volume = {32},
 year = {2019}
}
Defense Against Adversarial Attacks Using Feature Scattering-based Adversarial Training · NeurIPS 2019