arrow
返回

Explaining black-box classifiers: Properties and functions

delete2023-04-01
delete5
delete
OA
AI
L
Leïla Amgoud *
DOI:10.1016/j.ijar.2023.01.004delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Explaining black-box classification models is a hot topic in AI, with the overall goal of improving trust in decisions made by such models. Several works have been done and diverse functions have been proposed. However, their formal properties and links have not been sufficiently studied. This paper presents four contributions: The first consists of investigating global explanations of black-box classifiers. We provide a formal and unifying framework in which such explanations are defined from the whole feature space. The framework is based on two concepts, which are seen as two types of global explanations: arguments in favour of (or pro) predictions and arguments against (or con) predictions. The second contribution consists of defining various types of local explanations (abductive explanations, counterfactuals, contrastive explanations) from the whole feature space, investigating their properties, links and differences, and showing how they relate to global explanations. The third contribution consists of analysing and defining explanation functions that generate (global, local) abductive explanations from incomplete information (i.e., from a subset of the feature space). We start by proposing two desirable properties that an explainer would satisfy, namely success and coherence. The former ensures the existence of explanations while the latter ensures their correctness. We show that in the incomplete case, the two properties cannot be satisfied together. The fourth contribution consists of proposing two functions that generate abductive explanations and which satisfy coherence at the expense of success.(c) 2023 Elsevier Inc. All rights reserved.
Keyword:
Classification
Explainability
Arguments
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

International Journal of Approximate Reasoning 封面图
International Journal of Approximate Reasoning
IF:
3
论文数:
3.0K
被引数:
5.1K

机构

C
centre national de la recherche scientifique (cnrs)
学者数:
24.5W
论文数: 18.2W
被引数: 279
引用论文

引用论文

GLocalX - From Local to Global Explanations of Black Box AI ModelsGLocalX-从黑盒AI模型的局部到全局解释
err2021-05-01
err73
errOAAI
errSetzu, Mattia; Guidotti, Riccardo; Monreale, Anna; Turini, Franco; Pedreschi, Dino; Giannotti, Fosca
err分享
err收藏
The genome sequence of the rice blast fungus Magnaporthe grisea稻瘟病菌Magnaporthe grisea的基因组序列
err2005-04-01
err0
errOAAI
errRalph A. Dean; Nicholas J. Talbot; Daniel J. Ebbole; Mark L. Farman; Thomas K. Mitchell; Marc J. Orbach; Michael Thon; Resham Kulkarni; Jin-Rong Xu; Huaqin Pan; Nick D. Read; Yong-Hwan Lee; Ignazio Carbone; Doug Brown; Yeon Yee Oh; Nicole Donofrio; Jun Seop Jeong; Darren M. Soanes; Slavica Djonovic; Elena Kolomiets; Cathryn Rehmeyer; Weixi Li; Michael Harding; Soonok Kim; Marc-Henri Lebrun; Heidi Bohnert; Sean Coughlan; Jonathan Butler; Sarah Calvo; Li-Jun Ma; Robert Nicol; Seth Purcell; Chad Nusbaum; James E. Galagan; Bruce W. Birren
err分享
err收藏
Counterfactual Thought
err2016-01-04
err300
errOAAI
errByrne, Ruth M. J.
err分享
err收藏
err分享
err收藏
A Survey of Methods for Explaining Black Box Models黑箱模型的解释方法综述
err2018-08-22
err2.5K
errOAAI
errGuidotti, Riccardo; Monreale, Anna; Ruggieri, Salvatore; Turin, Franco; Giannotti, Fosca; Pedreschi, Dino
err分享
err收藏
学者 查看更多内容