Adversarial scratches: Deployable attacks to CNN classifiers
Abstract
A growing body of work has shown that deep neural networks are susceptible to adversarial examples. These take the form of small perturbations applied to the model's input which lead to incorrect predictions. Unfortunately, most literature focuses on visually imperceivable perturbations to be applied to digital images that often are, by design, impossible to be deployed to physical targets.
We present Adversarial Scratches: a novel L0 black-box attack, which takes the form of scratches in images, and which possesses much greater deployability than other state-of-the-art attacks. Adversarial Scratches leverage Bézier Curves to reduce the dimension of the search space and possibly constrain the attack to a specific location. We test Adversarial Scratches in several scenarios, including a publicly available API and images of traffic signs. Results show that our attack achieves higher fooling rate than other deployable state-of-the-art methods, while requiring significantly fewer queries and modifying very few pixels.- Publication:
-
Pattern Recognition
- Pub Date:
- January 2023
- DOI:
- 10.1016/j.patcog.2022.108985
- arXiv:
- arXiv:2204.09397
- Bibcode:
- 2023PatRe.13308985G
- Keywords:
-
- Adversarial perturbations;
- Adversarial attacks;
- Deep learning;
- Convolutional neural networks;
- Bézier curves;
- Computer Science - Machine Learning;
- Computer Science - Cryptography and Security;
- Computer Science - Computer Vision and Pattern Recognition;
- I.4;
- I.5
- E-Print:
- This work is published at Pattern Recognition (Elsevier). This paper stems from 'Scratch that! An Evolution-based Adversarial Attack against Neural Networks' for which an arXiv preprint is available at arXiv:1912.02316. Further studies led to a complete overhaul of the work, resulting in this paper