Adversarial Detection without Model Information

Moitra, Abhishek; Kim, Youngeun; Panda, Priyadarshini

Computer Science > Computer Vision and Pattern Recognition

arXiv:2202.04271 (cs)

[Submitted on 9 Feb 2022 (v1), last revised 5 Apr 2022 (this version, v2)]

Title:Adversarial Detection without Model Information

Authors:Abhishek Moitra, Youngeun Kim, Priyadarshini Panda

View PDF

Abstract:Prior state-of-the-art adversarial detection works are classifier model dependent, i.e., they require classifier model outputs and parameters for training the detector or during adversarial detection. This makes their detection approach classifier model specific. Furthermore, classifier model outputs and parameters might not always be accessible. To this end, we propose a classifier model independent adversarial detection method using a simple energy function to distinguish between adversarial and natural inputs. We train a standalone detector independent of the classifier model, with a layer-wise energy separation (LES) training to increase the separation between natural and adversarial energies. With this, we perform energy distribution-based adversarial detection. Our method achieves comparable performance with state-of-the-art detection works (ROC-AUC > 0.9) across a wide range of gradient, score and gaussian noise attacks on CIFAR10, CIFAR100 and TinyImagenet datasets. Furthermore, compared to prior works, our detection approach is light-weight, requires less amount of training data (40% of the actual dataset) and is transferable across different datasets. For reproducibility, we provide layer-wise energy separation training code at this https URL

Comments:	This paper has 14 pages of content and 2 pages of references
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2202.04271 [cs.CV]
	(or arXiv:2202.04271v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2202.04271

Submission history

From: Abhishek Moitra [view email]
[v1] Wed, 9 Feb 2022 04:38:16 UTC (1,029 KB)
[v2] Tue, 5 Apr 2022 05:21:58 UTC (2,272 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Adversarial Detection without Model Information

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Adversarial Detection without Model Information

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators