English

Empirical evaluation of the Frank-Wolfe methods for constructing white-box adversarial attacks

Machine Learning 2025-12-12 v1 Artificial Intelligence

Abstract

The construction of adversarial attacks for neural networks appears to be a crucial challenge for their deployment in various services. To estimate the adversarial robustness of a neural network, a fast and efficient approach is needed to construct adversarial attacks. Since the formalization of adversarial attack construction involves solving a specific optimization problem, we consider the problem of constructing an efficient and effective adversarial attack from a numerical optimization perspective. Specifically, we suggest utilizing advanced projection-free methods, known as modified Frank-Wolfe methods, to construct white-box adversarial attacks on the given input data. We perform a theoretical and numerical evaluation of these methods and compare them with standard approaches based on projection operations or geometrical intuition. Numerical experiments are performed on the MNIST and CIFAR-10 datasets, utilizing a multiclass logistic regression model, the convolutional neural networks (CNNs), and the Vision Transformer (ViT).

Keywords

Cite

@article{arxiv.2512.10936,
  title  = {Empirical evaluation of the Frank-Wolfe methods for constructing white-box adversarial attacks},
  author = {Kristina Korotkova and Aleksandr Katrutsa},
  journal= {arXiv preprint arXiv:2512.10936},
  year   = {2025}
}
R2 v1 2026-07-01T08:21:05.155Z