Return
On the Interplay of Explainability and Fairness in AI: A Survey
C
V
D
E
DOI:10.1109/tkde.2026.3706635.png)
Abstract
En 中文
Algorithmic fairness and explainability are foundational pillars of responsible AI. Although often studied independently, their interplay is increasingly recognized as crucial for diagnosing and mitigating bias in machine learning systems. We first introduce two systematic taxonomies: one for algorithmic fairness and one for explainable AI, to organize the landscape of existing work across diverse tasks (classification, ranking, and recommendation) and data modalities (tabular, graph). Next, we categorize the use of explanations in fairness efforts into three main functions: (a) detecting and understanding the causes of unfairness, (b) defining enhanced fairness metrics, and (c) designing mitigation strategies. In addition, we examine how explanation methods themselves can be biased, underscoring the need to evaluate fairness for explanations. Finally, we identify open research challenges and outline promising directions for future research at the intersection of fairness and explainability.
Keywords:
Fairness
explainability
responsible AI
Journal
IF:
10.4
Papers:
6.7K
Citations:
3.2W
