arrow
Return

Benchmarking Explainable AI Methods for Vision Transformer-Based Diabetic Retinopathy Analysis

delete2026-04-29
delete0
delete
OA
AI
L
Lucia Cascone
P
Pietro Campiglia
M
Michele Nappi
F
Fabio Narducci
B
Benedetto Simone
DOI:10.1109/ojcs.2026.3688710delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Vision Transformers have achieved notable results in diabetic retinopathy (DR) classification; yet, their interpretability remains a critical barrier to clinical adoption. Existing explainability studies often lack systematic comparisons across XAI paradigms and rely primarily on qualitative assessment. This work addresses this gap by presenting a comprehensive benchmark of six post-hoc explainability methods: Grad-CAM, Grad-CAM++, Score-CAM, Raw Attention, Attention Rollout, and AGCAM, applied to a fixed state-of-the-art Vision Transformer (Dino2-DR) using expert-annotated fundus images. Methods are evaluated along complementary dimensions, including lesion localization accuracy, intrinsic reliability (complexity, faithfulness, and robustness), and computational efficiency. Results indicate that no single method simultaneously optimizes all criteria. Grad-CAM achieves the strongest detection balance, Raw Attention provides superior spatial delineation, and AGCAM combines broad lesion coverage with the highest causal alignment to model predictions. Localization accuracy and intrinsic reliability emerge as partially independent quality dimensions, with gradient-based approaches favoring precision and hybrid methods emphasizing contextual coverage. Runtime analysis further highlights substantial efficiency differences across methods, identifying Grad-CAM and AGCAM as the most practical candidates for interactive deployment, while Score-CAM and Attention Rollout incur prohibitive overhead. These findings underscore the necessity of multi-dimensional explainability evaluation and provide actionable guidance for task-aware selection of XAI methods in Vision Transformer–based ophthalmic imaging.
Keywords:
Explainable AI
diabetic retinopathy
vision transformers
saliency maps
trustworthy medical AI
interpretability benchmark

Journal

I
IEEE Open Journal of the Computer Society
IF:
8.2
Papers:
411
Citations:
810

Organization

U
university of salerno
Scholars:
2.3K
Papers: 910
Citations: 0