返回
LDAS: Local density-based adaptive sampling for imbalanced data classification
DOI:10.1016/j.eswa.2021.116213.png)
摘要
En 中文
Class imbalance poses a great challenge to traditional classifiers in machine learning as they strongly favor the majority class while ignoring the minority class. Synthetic over-sampling methods deal with this problem by generating synthetic examples to balance the distribution of data. However, most existing methods prefer to generate synthetic examples in a specific area without considering the complexity of imbalance distribution, which may result in the over-emphasis of learning model on some data difficulty factors. To this end, we propose a local density-based adaptive sampling method (LDAS) for imbalanced data. LDAS first assigns a local density for each minority example, then a new cleaning strategy is proposed to remove the overlapping majority examples. Finally, it weighs each minority example based on its approaching degree of decision boundary and the corresponding local density. This is done in such a way that synthetic examples are generated in the safe area and the border area simultaneously according to the weight of minority examples. Extensive experiments on KEEL datasets demonstrate the effectiveness of the proposal LDAS.
Keyword:
Imbalanced classification
Local density
Overlapping data
Re-sampling
期刊
IF:
7.5
论文数:
2.9W
被引数:
10.2W
机构
引用论文
Radial-Based oversampling for noisy imbalanced data classification基于径向过采样的噪声不平衡数据分类
NEUROCOMPUTING
IF6.5
Arsenic removal from aqueous solutions by adsorption using novel MIL-53(Fe) as a highly efficient adsorbent使用新型MIL-53(Fe) 作为高效吸附剂通过吸附从水溶液中去除砷
RSC Advances
IF0

