Verifying Machine Unlearning with Explainable AI

Vidal, Àlex Pujol; Johansen, Anders S.; Jahromi, Mohammad N. S.; Escalera, Sergio; Nasrollahi, Kamal; Moeslund, Thomas B.

Computer Science > Machine Learning

arXiv:2411.13332 (cs)

[Submitted on 20 Nov 2024]

Title:Verifying Machine Unlearning with Explainable AI

Authors:Àlex Pujol Vidal, Anders S. Johansen, Mohammad N. S. Jahromi, Sergio Escalera, Kamal Nasrollahi, Thomas B. Moeslund

View PDF HTML (experimental)

Abstract:We investigate the effectiveness of Explainable AI (XAI) in verifying Machine Unlearning (MU) within the context of harbor front monitoring, focusing on data privacy and regulatory compliance. With the increasing need to adhere to privacy legislation such as the General Data Protection Regulation (GDPR), traditional methods of retraining ML models for data deletions prove impractical due to their complexity and resource demands. MU offers a solution by enabling models to selectively forget specific learned patterns without full retraining. We explore various removal techniques, including data relabeling, and model perturbation. Then, we leverage attribution-based XAI to discuss the effects of unlearning on model performance. Our proof-of-concept introduces feature importance as an innovative verification step for MU, expanding beyond traditional metrics and demonstrating techniques' ability to reduce reliance on undesired patterns. Additionally, we propose two novel XAI-based metrics, Heatmap Coverage (HC) and Attention Shift (AS), to evaluate the effectiveness of these methods. This approach not only highlights how XAI can complement MU by providing effective verification, but also sets the stage for future research to enhance their joint integration.

Comments:	ICPRW2024
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2411.13332 [cs.LG]
	(or arXiv:2411.13332v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2411.13332

Submission history

From: Alex Pujol Vidal [view email]
[v1] Wed, 20 Nov 2024 13:57:32 UTC (6,427 KB)

Computer Science > Machine Learning

Title:Verifying Machine Unlearning with Explainable AI

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Verifying Machine Unlearning with Explainable AI

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators