返回
Supervised discretization of continuous-valued attributes for classification using RACER algorithm
DOI:10.1016/j.eswa.2023.121203.png)
摘要
En 中文
In the contemporary world, data pervades every facet of human life, and the information contained in this data plays a pivotal role in shaping decision-making and advancing technology. Among the plethora of techniques available, classification methods are highly effective tools for extracting valuable insights from vast volumes of data. The Rule Aggregation ClassifiER (RACER) is a novel rule-based classification algorithm known for its exceptional performance. A notable limitation of RACER lies in its inability to handle continuous features. In this paper, we address the aforementioned limitation by employing various supervised discretization methods, including CAIM, MDLP, Decision Tree (CART), and ChiMerge. The impact of these methods on RACER's accuracy and understandability is evaluated across nine datasets from the UCI repository. Additionally, the paper conducts a comparative analysis of RACER's accuracy against well-known classifiers such as Naive Bayes, Logistic Regression, SVM, LightGBM, and Decision Tree. The findings indicate that RACER achieves the highest average accuracy when we utilize MDLP as the discretization method, surpassing its initial average accuracy. Moreover, RACER demonstrates superior understandability by generating the lowest number of rules when employing ChiMerge and Decision Tree for discretizing numerical features. Furthermore, RACER outperforms the other five classifiers when employing MDLP.
Keyword:
Rule aggregation classifiER (RACER)
Classification
Discretization
Rule-based classifier
期刊
IF:
7.5
论文数:
2.9W
被引数:
10.2W
机构
引用论文
The application of data mining techniques in financial fraud detection: A classification framework and an academic review of literature数据挖掘技术在财务舞弊检测中的应用: 一个分类框架和一个学术文献综述

