arrow
Return

CECE: Counterfactual Examples-based Causal Explanation Method

delete2025-10-01
delete0
PRE
AI
H
Haorun Ding
姜学松 (Xuesong Jiang) *
DOI:10.1007/s11760-025-04206-4delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Two common methods for interpreting machine learning models are counterfactual explanations and attributional explanations, each with its own advantages and limitations. Attributional interpretation assigns an importance score to each input feature, but it is difficult to ensure fidelity when interpreting complex models, and the commonly used attributional interpretations LIME and SHAP are based on the sufficiency definition. Counterfactual explanations provide minimally varying input instances to change model predictions, but are less likely to reflect generalizability, and their generated instances can be related to the necessity definition. To integrate these methods and leverage their respective advantages, we propose a novel approach for generating feature attribution explanations using counterfactual instances. We construct sufficiency impact terms and necessity impact terms from the generated counterfactual instances, and aggregate these terms by weight to derive feature importance scores. We evaluate it on three public datasets Adult-Income, Lending-Club, and German-Credit. The experimental results demonstrate that our method places greater emphasis on feature adequacy and necessity, and is more faithful to the original machine learning model.
Keywords:
Machine learning
Explanation
Feature attribution
Counterfactual

Journal

Signal Image and Video Processing cover
Signal Image and Video Processing
IF:
2.1
Papers:
877
Citations:
4.6K

Organization

Q
qilu university of technology
Scholars:
2.0K
Papers: 606
Citations: 0