arrow
返回

Deterministic Local Interpretable Model-Agnostic Explanations for Stable Explainability

delete2021-06-30
delete136
delete
OA
AI
M
Muhammad Rehman Zafar *
N
Naimul Khan
DOI:10.3390/make3030027delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Local Interpretable Model-Agnostic Explanations (LIME) is a popular technique used to increase the interpretability and explainability of black box Machine Learning (ML) algorithms. LIME typically creates an explanation for a single prediction by any ML model by learning a simpler interpretable model (e.g., linear classifier) around the prediction through generating simulated data around the instance by random perturbation, and obtaining feature importance through applying some form of feature selection. While LIME and similar local algorithms have gained popularity due to their simplicity, the random perturbation methods result in shifts in data and instability in the generated explanations, where for the same prediction, different explanations can be generated. These are critical issues that can prevent deployment of LIME in sensitive domains. We propose a deterministic version of LIME. Instead of random perturbation, we utilize Agglomerative Hierarchical Clustering (AHC) to group the training data together and K-Nearest Neighbour (KNN) to select the relevant cluster of the new instance that is being explained. After finding the relevant cluster, a simple model (i.e., linear model or decision tree) is trained over the selected cluster to generate the explanations. Experimental results on six public (three binary and three multi-class) and six synthetic datasets show the superiority for Deterministic Local Interpretable Model-Agnostic Explanations (DLIME), where we quantitatively determine the stability and faithfulness of DLIME compared to LIME.
Keyword:
explainable artificial intelligence (XAI)
interpretable machine learning
stable explanations
deterministic explanations
local explanations
model agnostic explanations
human interpretable explanations
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

M
Machine Learning and Knowledge Extraction
IF:
6
论文数:
818
被引数:
1.8K

机构

T
Toronto Metropolitan University
学者数:
6.0K
论文数: 7.0K
被引数: 6.4K
引用论文

引用论文

Reductions in neural activity underlie behavioral components of repetition priming
err2005-07-31
err0
PREAI
errGagan S Wig; Scott T Grafton; Kathryn E Demos; William M Kelley
err分享
err收藏
err分享
err收藏
err分享
err收藏
High-energy neutrino fluxes and flavor ratio in the Earth’s atmosphere
err2015-03-31
err0
errOAAI
errT. S. Sinegovskaya; A. D. Morozova; S. I. Sinegovsky
err分享
err收藏
err分享
err收藏
err分享
err收藏
Factual and Counterfactual Explanations for Black Box Decision Making黑箱决策的事实与反事实解释
err2019-11-01
err169
errOAAI
errGuidotti, Riccardo; Monreale, Anna; Giannotti, Fosca; Pedreschi, Dino; Ruggieri, Salvatore; Turini, Franco
err分享
err收藏
学者 查看更多内容