A survey of methods for explaining black box models

Обзор методов объяснения «черных ящиков»
Dino Pedreschi, Riccardo Guidotti, Anna Monreale, Salvatore Ruggieri, Franco Turini, Fosca Giannotti, Salvatore Ruggieri
2019-01-01

black box modelsexplainable AIinterpretability taxonomymodel interpretabilitypost-hoc explanation methods
In recent years, many accurate decision support systems have been constructed as black boxes, that is as systems that hide their internal logic to the user. This lack of explanation constitutes both a practical and an ethical issue. The literature reports many approaches aimed at overcoming this crucial weakness, sometimes at the cost of sacrificing accuracy for interpretability. The applications in which black box decision systems can be used are various, and each approach is typically developed to provide a solution for a specific problem and, as a consequence, it explicitly or implicitly delineates its own definition of interpretability and explanation. The aim of this article is to provide a classification of the main problems addressed in the literature with respect to the notion of explanation and the type of black box system. Given a problem definition, a black box type, and a desired explanation, this survey should help the researcher to find the proposals more useful for his own work. The proposed classification of approaches to open black box models should also be useful for putting the many research open questions in perspective.
1
Black box decision support systems are widespread and their lack of explanations poses practical and ethical problems.
2
Different application domains and problem settings lead to diverse, sometimes implicit, definitions of interpretability and explanation.
3
Many approaches exist to make black boxes explainable, often trading accuracy for interpretability.
4
The article provides a classification of main problems, black box types, and desired explanations to guide method selection.
5
The proposed classification organizes existing approaches and highlights open research questions in explaining black box models.

Black-box decision support systems (black box models)

Methods for explaining/opening black-box models: classification of approaches, definitions of interpretability and explanation, and mapping techniques to problem types and black-box types

Publication Details
Publication Date
2019-01-01
Journal
Publisher
ISSN
Access Type
Author Information
Authors
Dino Pedreschi
Riccardo Guidotti
Anna Monreale
Salvatore Ruggieri
Franco Turini
Fosca Giannotti
Salvatore Ruggieri
Explore further
Open the scid.ai AI chat with a ready-made request: it will find papers on a similar topic and help build a literature review.
Find similar papers in the chat
Make a presentation
100%