Mathematical Analysis of Adversarial Attacks

Bao Wang; Stanley J. Osher; Zehao Dou

arxiv: 1811.06492 · v2 · pith:Y5UTCFQEnew · submitted 2018-11-15 · 💻 cs.LG · cs.CR· stat.ML

Mathematical Analysis of Adversarial Attacks

Zehao Dou , Stanley J. Osher , Bao Wang This is my paper

classification 💻 cs.LG cs.CRstat.ML

keywords activationfgsmattackcnnscw-l2neuralreluresults

0 comments

read the original abstract

In this paper, we analyze efficacy of the fast gradient sign method (FGSM) and the Carlini-Wagner's L2 (CW-L2) attack. We prove that, within a certain regime, the untargeted FGSM can fool any convolutional neural nets (CNNs) with ReLU activation; the targeted FGSM can mislead any CNNs with ReLU activation to classify any given image into any prescribed class. For a special two-layer neural network: a linear layer followed by the softmax output activation, we show that the CW-L2 attack increases the ratio of the classification probability between the target and ground truth classes. Moreover, we provide numerical results to verify all our theoretical results.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

Graph Interpolating Activation Improves Both Natural and Robust Accuracies in Data-Efficient Deep Learning
cs.LG 2019-07 unverdicted novelty 5.0

Graph Laplacian interpolating activation replaces softmax in DNNs and improves natural accuracy, robust accuracy, and data efficiency.