Mathematical Analysis of Adversarial Attacks

Dou, Zehao, Osher, Stanley J., Wang, Bao

Nov-15-2018–arXiv.org Machine Learning

In this paper, we analyze efficacy of the fast gradient sign method (FGSM) and the Carlini-Wagner's L2 (CW-L2) attack. We prove that, within a certain regime, the untargeted FGSM can fool any convolutional neural nets (CNNs) with ReLU activation; the targeted FGSM can mislead any CNNs with ReLU activation to classify any given image into any prescribed class. For a special two-layer neural network: a linear layer followed by the softmax output activation, we show that the CW-L2 attack increases the ratio of the classification probability between the target and ground truth classes. Moreover, we provide numerical results to verify all our theoretical results.

deep learning, exp, neural network, (17 more...)

arXiv.org Machine Learning

Nov-15-2018

arXiv.org PDF

Add feedback

Country:
- North America > United States (0.93)

Genre:
- Research Report (0.64)

Industry:
- Government
  - Military (0.86)
  - Regional Government > North America Government
    - United States Government (0.68)
- Information Technology > Security & Privacy (0.87)

Technology:
- Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (0.46)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found