arrow
Return

A generalized model for predictive data mining

delete2002-01-01
delete6
PRE
AI
J
James V. Hansen *
J
James B. McDonald
DOI:10.1023/A:1016050803099delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
This paper describes a flexible model for predictive data mining, EGB2, which optimizes over a parameter space to fit data to a family of models based on maximum-likelihood criteria. It is also shown how EGB2 can integrate asymmetric costs of Type I and Type II errors, thereby minimizing expected misclassification costs. Importantly, it has been shown that standard methods of computing maximum-likelihood estimators are generally inconsistent when applied to sample data having different proportions of labels than are found in the universe from which the sample is drawn. We show how a choice estimator based on weighting each observation's contribution to the log-likelihood function, can contribute to estimator consistency and how this feature can be implemented in EGB2.
Keywords:
data mining
prediction
choice estimator
misclassification costs
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Information Systems Frontiers cover
Information Systems Frontiers
IF:
8.3
Papers:
2.0K
Citations:
6.5K

Organization

No organization information available