arrow
返回

Data Mining for Discrimination Discovery

delete2010-05-28
delete81
delete
OA
AI
S
Salvatore Ruggieri *
D
Dino Pedreschi
F
Franco Turini
DOI:10.1145/1754428.1754432delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
In the context of civil rights law, discrimination refers to unfair or unequal treatment of people based on membership to a category or a minority, without regard to individual merit. Discrimination in credit, mortgage, insurance, labor market, and education has been investigated by researchers in economics and human sciences. With the advent of automatic decision support systems, such as credit scoring systems, the ease of data collection opens several challenges to data analysts for the fight against discrimination. In this article, we introduce the problem of discovering discrimination through data mining in a dataset of historical decision records, taken by humans or by automatic systems. We formalize the processes of direct and indirect discrimination discovery by modelling protected-by-law groups and contexts where discrimination occurs in a classification rule based syntax. Basically, classification rules extracted from the dataset allow for unveiling contexts of unlawful discrimination, where the degree of burden over protected-by-law groups is formalized by an extension of the lift measure of a classification rule. In direct discrimination, the extracted rules can be directly mined in search of discriminatory contexts. In indirect discrimination, the mining process needs some background knowledge as a further input, for example, census data, that combined with the extracted rules might allow for unveiling contexts of discriminatory decisions. A strategy adopted for combining extracted classification rules with background knowledge is called an inference model. In this article, we propose two inference models and provide automatic procedures for their implementation. An empirical assessment of our results is provided on the German credit dataset and on the PKDD Discovery Challenge 1999 financial dataset.
Keyword:
Discrimination
classification rules
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

ACM Transactions on Knowledge Discovery from Data 封面图
ACM Transactions on Knowledge Discovery from Data
IF:
4.8
论文数:
1.3K
被引数:
4.4K

机构

U
University of Pisa
学者数:
3.1W
论文数: 2.4W
被引数: 2.4W
引用论文

引用论文

WHIPPLEʼS DISEASE
err1970-05-01
err0
errOAAI
errHAROLD MAIZEL; JULIAN M. RUFFIN; WILLIAM O. DOBBINS
err分享
err收藏
err分享
err收藏
学者 查看更多内容