Investigating Adversarial Vulnerability and Implicit Bias through Frequency Analysis

Basile, Lorenzo; Karantzas, Nikos; D'Onofrio, Alberto; Bortolussi, Luca; Rodriguez, Alex; Anselmi, Fabio

Computer Science > Machine Learning

arXiv:2305.15203 (cs)

[Submitted on 24 May 2023 (v1), last revised 17 Jul 2024 (this version, v2)]

Title:Investigating Adversarial Vulnerability and Implicit Bias through Frequency Analysis

Authors:Lorenzo Basile, Nikos Karantzas, Alberto D'Onofrio, Luca Bortolussi, Alex Rodriguez, Fabio Anselmi

View PDF HTML (experimental)

Abstract:Despite their impressive performance in classification tasks, neural networks are known to be vulnerable to adversarial attacks, subtle perturbations of the input data designed to deceive the model. In this work, we investigate the relation between these perturbations and the implicit bias of neural networks trained with gradient-based algorithms. To this end, we analyse the network's implicit bias through the lens of the Fourier transform. Specifically, we identify the minimal and most critical frequencies necessary for accurate classification or misclassification respectively for each input image and its adversarially perturbed version, and uncover the correlation among those. To this end, among other methods, we use a newly introduced technique capable of detecting non-linear correlations between high-dimensional datasets. Our results provide empirical evidence that the network bias in Fourier space and the target frequencies of adversarial attacks are highly correlated and suggest new potential strategies for adversarial defence.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Cryptography and Security (cs.CR); Machine Learning (stat.ML)
Cite as:	arXiv:2305.15203 [cs.LG]
	(or arXiv:2305.15203v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2305.15203

Submission history

From: Lorenzo Basile [view email]
[v1] Wed, 24 May 2023 14:40:23 UTC (6,763 KB)
[v2] Wed, 17 Jul 2024 16:34:48 UTC (3,445 KB)

Computer Science > Machine Learning

Title:Investigating Adversarial Vulnerability and Implicit Bias through Frequency Analysis

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Investigating Adversarial Vulnerability and Implicit Bias through Frequency Analysis

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators