Turning Your Strength against You: Detecting and Mitigating Robust and Universal Adversarial Patch Attacks

Chen, Zitao; Dash, Pritam; Pattabiraman, Karthik

Computer Science > Cryptography and Security

arXiv:2108.05075v3 (cs)

[Submitted on 11 Aug 2021 (v1), revised 23 Jun 2022 (this version, v3), latest version 16 Dec 2022 (v4)]

Title:Turning Your Strength against You: Detecting and Mitigating Robust and Universal Adversarial Patch Attacks

Authors:Zitao Chen, Pritam Dash, Karthik Pattabiraman

View PDF

Abstract:Adversarial patch attacks that inject arbitrary distortions within a bounded region of an image, can trigger misclassification in deep neural networks (DNNs). These attacks are robust (i.e., physically realizable) and universally malicious, and hence represent a severe security threat to real-world DNN-based systems.
This work proposes Jujutsu, a two-stage technique to detect and mitigate robust and universal adversarial patch attacks. We first observe that patch attacks often yield large influence on the prediction output in order to dominate the prediction on any input, and Jujutsu is built to expose this behavior for effective attack detection. For mitigation, we observe that patch attacks corrupt only a localized region while the remaining contents are unperturbed, based on which Jujutsu leverages GAN-based image inpainting to synthesize the semantic contents in the pixels that are corrupted by the attacks, and reconstruct the ``clean'' image for correct prediction.
We evaluate Jujutsu on four diverse datasets and show that it achieves superior performance and significantly outperforms four leading defenses. Jujutsu can further defend against physical-world attacks, attacks that target diverse classes, and adaptive attacks. Our code is available at this https URL.

Subjects:	Cryptography and Security (cs.CR); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2108.05075 [cs.CR]
	(or arXiv:2108.05075v3 [cs.CR] for this version)
	https://doi.org/10.48550/arXiv.2108.05075

Submission history

From: Zitao Chen [view email]
[v1] Wed, 11 Aug 2021 07:37:03 UTC (4,263 KB)
[v2] Wed, 17 Nov 2021 03:00:58 UTC (8,393 KB)
[v3] Thu, 23 Jun 2022 19:04:28 UTC (8,062 KB)
[v4] Fri, 16 Dec 2022 08:49:08 UTC (2,148 KB)

Computer Science > Cryptography and Security

Title:Turning Your Strength against You: Detecting and Mitigating Robust and Universal Adversarial Patch Attacks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Cryptography and Security

Title:Turning Your Strength against You: Detecting and Mitigating Robust and Universal Adversarial Patch Attacks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators