arrow
Return

LDAS: Local density-based adaptive sampling for imbalanced data classification

delete2022-04-01
delete33
PRE
AI
严远亭 (Yuanting Yan) *
Y
Yifei Jiang
Z
Zhong Zheng
C
Chengjin Yu
张议文 (Yiwen Zhang)
张艳平 cover
张艳平 (Yanping Zhang)
DOI:10.1016/j.eswa.2021.116213delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Class imbalance poses a great challenge to traditional classifiers in machine learning as they strongly favor the majority class while ignoring the minority class. Synthetic over-sampling methods deal with this problem by generating synthetic examples to balance the distribution of data. However, most existing methods prefer to generate synthetic examples in a specific area without considering the complexity of imbalance distribution, which may result in the over-emphasis of learning model on some data difficulty factors. To this end, we propose a local density-based adaptive sampling method (LDAS) for imbalanced data. LDAS first assigns a local density for each minority example, then a new cleaning strategy is proposed to remove the overlapping majority examples. Finally, it weighs each minority example based on its approaching degree of decision boundary and the corresponding local density. This is done in such a way that synthetic examples are generated in the safe area and the border area simultaneously according to the weight of minority examples. Extensive experiments on KEEL datasets demonstrate the effectiveness of the proposal LDAS.
Keywords:
Imbalanced classification
Local density
Overlapping data
Re-sampling

Journal

Expert Systems with Applications cover
Expert Systems with Applications
IF:
7.5
Papers:
2.9W
Citations:
10.2W

Organization

Z
zhejiang university
Scholars:
17.6W
Papers: 12.1W
Citations: 152
A
anhui university
Scholars:
1.9W
Papers: 1.2W
Citations: 24