English

Adversarial Robustness against Multiple and Single $l_p$-Threat Models via Quick Fine-Tuning of Robust Classifiers

Machine Learning 2022-08-09 v2 Cryptography and Security Computer Vision and Pattern Recognition

Abstract

A major drawback of adversarially robust models, in particular for large scale datasets like ImageNet, is the extremely long training time compared to standard ones. Moreover, models should be robust not only to one lpl_p-threat model but ideally to all of them. In this paper we propose Extreme norm Adversarial Training (E-AT) for multiple-norm robustness which is based on geometric properties of lpl_p-balls. E-AT costs up to three times less than other adversarial training methods for multiple-norm robustness. Using E-AT we show that for ImageNet a single epoch and for CIFAR-10 three epochs are sufficient to turn any lpl_p-robust model into a multiple-norm robust model. In this way we get the first multiple-norm robust model for ImageNet and boost the state-of-the-art for multiple-norm robustness to more than 51%51\% on CIFAR-10. Finally, we study the general transfer via fine-tuning of adversarial robustness between different individual lpl_p-threat models and improve the previous SOTA l1l_1-robustness on both CIFAR-10 and ImageNet. Extensive experiments show that our scheme works across datasets and architectures including vision transformers.

Keywords

Cite

@article{arxiv.2105.12508,
  title  = {Adversarial Robustness against Multiple and Single $l_p$-Threat Models via Quick Fine-Tuning of Robust Classifiers},
  author = {Francesco Croce and Matthias Hein},
  journal= {arXiv preprint arXiv:2105.12508},
  year   = {2022}
}

Comments

ICML 2022