arrow
Return

Classification algorithm sensitivity to training data with non representative attribute noise

delete2009-02-01
delete30
PRE
AI
M
Michael V. Mannino *
Y
Yanjuan Yang
Y
Young U. Ryu
DOI:10.1016/j.dss.2008.11.021delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
We present an empirical comparison of classification algorithms when training data contains attribute noise levels not representative of field data. To study algorithm sensitivity, we develop an innovative experimental design using noise situation, algorithm, noise level, and training set size as factors. Our results contradict conventional wisdom indicating that investments to achieve representative noise levels may not be worthwhile. ill general, over representative training noise Should be avoided while under representative training noise is less of a concern. However, interactions among algorithm, noise level, and training set size indicate that these general results may not apply to particular practice situations. (c) 2008 Elsevier B.V. All rights reserved.
Keywords:
Attribute noise
Area under the Receiver Operating Curve
Classification algorithm

Journal

Decision Support Systems cover
Decision Support Systems
IF:
6.8
Papers:
3.8K
Citations:
1.5W

Organization

University of Colorado System cover
University of Colorado System
Scholars:
6.3W
Papers: 5.5W
Citations: 1.8K
U
University of Colorado Denver
Scholars:
5.3K
Papers: 4.2K
Citations: 10