A Survey Of Methods For Explaining Black Box Models

Guidotti, Riccardo; Monreale, Anna; Turini, Franco; Pedreschi, Dino; Giannotti, Fosca

Computer Science > Computers and Society

arXiv:1802.01933v2 (cs)

[Submitted on 6 Feb 2018 (v1), revised 19 Feb 2018 (this version, v2), latest version 21 Jun 2018 (v3)]

Title:A Survey Of Methods For Explaining Black Box Models

Authors:Riccardo Guidotti, Anna Monreale, Franco Turini, Dino Pedreschi, Fosca Giannotti

View PDF

Abstract:In the last years many accurate decision support systems have been constructed as black boxes, that is as systems that hide their internal logic to the user. This lack of explanation constitutes both a practical and an ethical issue. The literature reports many approaches aimed at overcoming this crucial weakness sometimes at the cost of scarifying accuracy for interpretability. The applications in which black box decision systems can be used are various, and each approach is typically developed to provide a solution for a specific problem and, as a consequence, delineating explicitly or implicitly its own definition of interpretability and explanation. The aim of this paper is to provide a classification of the main problems addressed in the literature with respect to the notion of explanation and the type of black box system. Given a problem definition, a black box type, and a desired explanation this survey should help the researcher to find the proposals more useful for his own work. The proposed classification of approaches to open black box models should also be useful for putting the many research open questions in perspective.

Comments:	This work is currently under review on an international journal
Subjects:	Computers and Society (cs.CY); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:1802.01933 [cs.CY]
	(or arXiv:1802.01933v2 [cs.CY] for this version)
	https://doi.org/10.48550/arXiv.1802.01933

Submission history

From: Riccardo Guidotti [view email]
[v1] Tue, 6 Feb 2018 13:20:02 UTC (1,512 KB)
[v2] Mon, 19 Feb 2018 12:29:56 UTC (1,512 KB)
[v3] Thu, 21 Jun 2018 08:15:38 UTC (1,512 KB)

Computer Science > Computers and Society

Title:A Survey Of Methods For Explaining Black Box Models

Submission history

Access Paper:

References & Citations

2 blog links

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computers and Society

Title:A Survey Of Methods For Explaining Black Box Models

Submission history

Access Paper:

References & Citations

2 blog links

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators