arrow
Return

Selective oversampling approach for strongly imbalanced data

delete2021-06-18
delete40
delete
OA
AI
P
Peter Gnip
L
Liberios Vokorokos
P
Peter Drotár *
DOI:10.7717/peerj-cs.604delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Challenges posed by imbalanced data are encountered in many real-world applications. One of the possible approaches to improve the classifier performance on imbalanced data is oversampling. In this paper, we propose the new selective oversampling approach (SOA) that first isolates the most representative samples from minority classes by using an outlier detection technique and then utilizes these samples for synthetic oversampling. We show that the proposed approach improves the performance of two state-of-the-art oversampling methods, namely, the synthetic minority oversampling technique and adaptive synthetic sampling. The prediction performance is evaluated on four synthetic datasets and four real-world datasets, and the proposed SOA methods always achieved the same or better performance than other considered existing oversampling methods.
Keywords:
Imbalanced data
Oversampling
Outlier detection
SMOTE
ADASYN
Bankruptcy prediction
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

PeerJ Computer Science cover
PeerJ Computer Science
IF:
2.5
Papers:
3.4K
Citations:
6.9K

Organization

T
technical university kosice
Scholars:
2.7K
Papers: 1.8K
Citations: 6