Ground Truth Evaluation of Neural Network Explanations with CLEVR-XAI

Arras, Leila; Osman, Ahmed; Samek, Wojciech

doi:10.1016/j.inffus.2021.11.008

Computer Science > Computer Vision and Pattern Recognition

arXiv:2003.07258 (cs)

[Submitted on 16 Mar 2020 (v1), last revised 9 Feb 2021 (this version, v2)]

Title:Ground Truth Evaluation of Neural Network Explanations with CLEVR-XAI

Authors:Leila Arras, Ahmed Osman, Wojciech Samek

View PDF

Abstract:The rise of deep learning in today's applications entailed an increasing need in explaining the model's decisions beyond prediction performances in order to foster trust and accountability. Recently, the field of explainable AI (XAI) has developed methods that provide such explanations for already trained neural networks. In computer vision tasks such explanations, termed heatmaps, visualize the contributions of individual pixels to the prediction. So far XAI methods along with their heatmaps were mainly validated qualitatively via human-based assessment, or evaluated through auxiliary proxy tasks such as pixel perturbation, weak object localization or randomization tests. Due to the lack of an objective and commonly accepted quality measure for heatmaps, it was debatable which XAI method performs best and whether explanations can be trusted at all. In the present work, we tackle the problem by proposing a ground truth based evaluation framework for XAI methods based on the CLEVR visual question answering task. Our framework provides a (1) selective, (2) controlled and (3) realistic testbed for the evaluation of neural network explanations. We compare ten different explanation methods, resulting in new insights about the quality and properties of XAI methods, sometimes contradicting with conclusions from previous comparative studies. The CLEVR-XAI dataset and the benchmarking code can be found at this https URL.

Comments:	37 pages, 9 tables, 2 figures (plus appendix 14 pages)
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Neural and Evolutionary Computing (cs.NE); Image and Video Processing (eess.IV)
Cite as:	arXiv:2003.07258 [cs.CV]
	(or arXiv:2003.07258v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2003.07258
Related DOI:	https://doi.org/10.1016/j.inffus.2021.11.008

Submission history

From: Wojciech Samek [view email]
[v1] Mon, 16 Mar 2020 14:43:33 UTC (1,114 KB)
[v2] Tue, 9 Feb 2021 16:18:05 UTC (3,542 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Ground Truth Evaluation of Neural Network Explanations with CLEVR-XAI

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Ground Truth Evaluation of Neural Network Explanations with CLEVR-XAI

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators