arrow
Return

Interpretable and explainable machine learning: A methods-centric overview with concrete examples

delete2023-02-28
delete49
delete
OA
AI
R
Ričards Marcinkevičs *
J
Julia E. Vogt
DOI:10.1002/widm.1493delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Interpretability and explainability are crucial for machine learning (ML) and statistical applications in medicine, economics, law, and natural sciences and form an essential principle for ML model design and development. Although interpretability and explainability have escaped a precise and universal definition, many models and techniques motivated by these properties have been developed over the last 30 years, with the focus currently shifting toward deep learning. We will consider concrete examples of state-of-the-art, including specially tailored rule-based, sparse, and additive classification models, interpretable representation learning, and methods for explaining black-box models post hoc. The discussion will emphasize the need for and relevance of interpretability and explainability, the divide between them, and the inductive biases behind the presented zoo of interpretable models and explanation methods.This article is categorized under:Fundamental Concepts of Data and Knowledge > Explainable AITechnologies > Machine LearningCommercial, Legal, and Ethical Issues > Social Considerations
Keywords:
explainability
interpretability
machine learning
neural networks
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Wiley Interdisciplinary Reviews-Data Mining and Knowledge Discovery cover
Wiley Interdisciplinary Reviews-Data Mining and Knowledge Discovery
IF:
11.7
Papers:
533
Citations:
5.3K

Organization

S
swiss federal institutes of technology domain
Scholars:
9.0W
Papers: 8.0W
Citations: 163