arrow
Return

Filter-Based Data Partitioning for Training Multiple Classifier Systems

delete2010-04-01
delete7
PRE
AI
R
Rozita Dara *
M
Masoud Makrehchi
M
Mohamed S. Kamel
DOI:10.1109/TKDE.2009.80delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Data partitioning methods such as bagging and boosting have been extensively used in multiple classifier systems. These methods have shown a great potential for improving classification accuracy. This study is concerned with the analysis of training data distribution and its impact on the performance of multiple classifier systems. In this study, several feature-based and class-based measures are proposed. These measures can be used to estimate statistical characteristics of the training partitions. To assess the effectiveness of different types of training partitions, we generated a large number of disjoint training partitions with distinctive distributions. Then, we empirically assessed these training partitions and their impact on the performance of the system by utilizing the proposed feature-based and class-based measures. We applied the findings of this analysis and developed a new partitioning method called Clustering, Declustering, and Selection (CDS). This study presents a comparative analysis of several existing data partitioning methods including our proposed CDS approach.
Keywords:
Multiple classifier system
combining method
wrapper-based data partitioning
filter-based data partitioning
distance
feature-based
class-based
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Transactions on Knowledge and Data Engineering cover
IEEE Transactions on Knowledge and Data Engineering
IF:
10.4
Papers:
6.7K
Citations:
3.2W

Organization

U
University of Waterloo
Scholars:
2.2W
Papers: 2.3W
Citations: 3.3W
C
clarivate
Scholars:
395
Papers: 300
Citations: 0