SparseFool: a few pixels make a big difference (1811.02248v4)

Published 6 Nov 2018 in cs.CV, cs.CR, and cs.LG

Abstract: Deep Neural Networks have achieved extraordinary results on image classification tasks, but have been shown to be vulnerable to attacks with carefully crafted perturbations of the input data. Although most attacks usually change values of many image's pixels, it has been shown that deep networks are also vulnerable to sparse alterations of the input. However, no computationally efficient method has been proposed to compute sparse perturbations. In this paper, we exploit the low mean curvature of the decision boundary, and propose SparseFool, a geometry inspired sparse attack that controls the sparsity of the perturbations. Extensive evaluations show that our approach computes sparse perturbations very fast, and scales efficiently to high dimensional data. We further analyze the transferability and the visual effects of the perturbations, and show the existence of shared semantic information across the images and the networks. Finally, we show that adversarial training can only slightly improve the robustness against sparse additive perturbations computed with SparseFool.

Authors (3)

Apostolos Modas (13 papers)
Seyed-Mohsen Moosavi-Dezfooli (33 papers)
Pascal Frossard (194 papers)

Citations (186)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

SparseFool: a few pixels make a big difference (1811.02248v4)

Summary

Related Papers