Explainable AI Applications in Healthcare: A Systematic Review

Применение объяснимого искусственного интеллекта в здравоохранении: систематический обзор
Ojobo Agbo Eje, Sayed Mehedi Azim, Abdollah Dehzangi
2026-06-17

Explainability evaluation metricsExplainable artificial intelligenceGrad-CAM and LRPHealthcareSHAP and LIME
Artificial Intelligence (AI) shows significant potential across healthcare domains, including advanced diagnostics, clinical decision support, and personalized medicine. Despite these advancements, the opaque ‘black box’ nature of complex AI models necessitates the application of Explainable Artificial Intelligence (XAI) to ensure trust, accountability, interpretability, and regulatory compliance. This study systematically reviews 76 studies published between 2020 and 2025 that have used XAI in healthcare. Our findings show that XAI models such as SHAP and LIME are predominantly used for structured data applications, such as electronic health records, while other XAI models, such as Grad-CAM and Layer-wise Relevance Propagation (LRP), are mainly used in medical imaging. This study specifically investigates evaluation metrics for operationalizing explainability, including faithfulness, trustworthiness, and regulatory compliance, which distinguishes it from prior descriptive reviews. Our analysis shows that while XAI significantly enhances clinician trust, thorough explanation remains heterogeneous and largely confined to controlled settings and the employed benchmark datasets. Critical barriers to clinical adoption include inconsistent interpretability across data modalities and the lack of standardized evaluation frameworks. Existing XAI techniques often do not correspond with strict regulatory standards such as the EU AI Act, Food and Drug Administration (FDA) guidelines, and the Health Insurance Portability and Accountability Act (HIPAA). This review argues for the urgent standardization of XAI validation and the development of human-centered designs to move beyond algorithmic transparency toward reliable real-world hospital integration.
1
Clinical adoption is hindered by inconsistent interpretability across modalities, absent standardized evaluation frameworks, and limited alignment with EU AI Act, FDA, and HIPAA requirements.
2
SHAP and LIME predominantly support structured-data applications such as electronic health records, whereas Grad-CAM and LRP are mainly used for medical imaging.
3
The review evaluates explainability through faithfulness, trustworthiness, and regulatory compliance, extending beyond prior descriptive reviews.
4
The systematic review analyzes 76 healthcare studies published between 2020 and 2025 that applied explainable artificial intelligence methods.
5
XAI improves clinician trust, but explanation quality remains heterogeneous and is largely demonstrated only in controlled settings and benchmark datasets.

Explainable Artificial Intelligence (XAI) applications in healthcare, including electronic health records and medical imaging

Operational evaluation of XAI explainability in healthcare, focusing on faithfulness, trustworthiness, regulatory compliance, clinician trust, interpretability across data modalities, and clinical adoption

Publication Details
Publication Date
2026-06-17
Journal
Publisher
ISSN
Cited by
3
Access Type
Author Information
Authors
Ojobo Agbo Eje
Sayed Mehedi Azim
Abdollah Dehzangi
Explore further
Open the scid.ai AI chat with a ready-made request: it will find papers on a similar topic and help build a literature review.
Find similar papers in the chat
Make a presentation
100%