返回
Dealing with class imbalance in classifier chains via random undersampling
DOI:10.1016/j.knosys.2019.105292.png)
摘要
En 中文
Class imbalance is an intrinsic characteristic of multi-label data. Most of the labels in multi-label data sets are associated with a small number of training examples, much smaller compared to the size of the data set. Class imbalance poses a key challenge that plagues most multi-label learning methods. Ensemble of Classifier Chains (ECC), one of the most prominent multi-label learning methods, is no exception to this rule, as each of the binary models it builds is trained from all positive and negative examples of a label. To make ECC resilient to class imbalance, we first couple it with random undersampling. We then present two extensions of this basic approach, where we build a varying number of binary models per label and construct chains of different sizes, in order to improve the exploitation of majority examples with approximately the same computational budget. Experimental results on 16 multi-label datasets demonstrate the effectiveness of the proposed approaches in a variety of evaluation metrics. (C) 2019 Elsevier B.V. All rights reserved.
Keyword:
Multi-label learning
Class imbalance
Classifier chains
Undersampling
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
K
IF:
7.6
论文数:
1.3W
被引数:
4.5W
机构
引用论文
HPSLPred: An Ensemble Multi-Label Classifier for Human Protein Subcellular Location Prediction with Imbalanced Source
PROTEOMICS
IF3.9
Addressing imbalance in multilabel classification: Measures and random resampling algorithms
NEUROCOMPUTING
IF6.5
Inverse random under sampling for class imbalance problem and its application to multi-label classification
PATTERN RECOGNITION
IF7.6

