arrow
返回

Affinity and class probability-based fuzzy support vector machine for imbalanced data sets

delete2020-02-01
delete50
PRE
AI
X
Xinmin Tao *
Prof. LI Qing 封面图
Prof. LI Qing (Qing Li)
C
Chao Ren
W
Wenjie Guo
Q
Qing He
R
Rui Liu
J
Junrong Zou
DOI:10.1016/j.neunet.2019.10.016delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The learning problem from imbalanced data sets poses a major challenge in data mining community. Although conventional support vector machine can generally show relatively robust performance in dealing with the classification problems of imbalanced data sets, it treats all training samples with the same contribution for learning, which results in the final decision boundary biasing toward the majority class especially in the presence of outliers or noises. In this paper, we propose a new affinity and class probability-based fuzzy support vector machine technique (ACFSVM). The affinity of a majority class sample is calculated according to support vector description domain (SVDD) model trained only by the given majority class training samples in kernel space similar to that used for FSVM learning. The obtained affinity can be used for identifying possible outliers and some border samples existing in the majority class training samples. In order to eliminate the effect of noises, we employ the kernel k-nearest neighbor method to determine the class probability of the majority class samples in the same kernel space as before. The samples with lower class probabilities are more likely to be noises and their contribution for learning seems to be reduced by their low memberships constructed by combining the affinities and the class probabilities. Thus, ACFSVM can pay more attention to the majority class samples with higher affinities and class probabilities while reducing their effects of the ones with lower affinities and class probabilities, eventually skewing the final classification boundary toward the majority class. In addition, the minority class samples are assigned relative high memberships to guarantee their importance for the model learning. The extensive experimental results on the different imbalanced datasets from UCI repository demonstrate that the proposed approach can achieve better generalization performance in terms of G-Mean, F-Measure, and AUC as compared to the other existing imbalanced dataset classification techniques. (c) 2019 Elsevier Ltd. All rights reserved.
Keyword:
Imbalanced data
Fuzzy support vector machine (FSVM)
Affinity
Class probability
Kernelknn
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Neural Networks 封面图
Neural Networks
IF:
6.3
论文数:
8.1K
被引数:
3.0W

机构

N
northeast forestry university - china
学者数:
1.2W
论文数: 7.9K
被引数: 9
引用论文

引用论文

Hepatic Surgery
err
IF0
err2013-02-13
err0
PREAI
err
err分享
err收藏
err
IF0
err
err0
PREAI
err
err分享
err收藏
Face memory and face recognition in children and adolescents with attention deficit hyperactivity disorder: A systematic review
err2018-06-01
err30
PREAI
errRomani, Maria; Vigliante, Miriam; Faedda, Noemi; Rossetti, Serena; Pezzuti, Lina; Guidetti, Vincenzo; Cardona, Francesco
err分享
err收藏
err分享
err收藏
Arsenic removal from aqueous solutions by adsorption using novel MIL-53(Fe) as a highly efficient adsorbent使用新型MIL-53(Fe) 作为高效吸附剂通过吸附从水溶液中去除砷
err2015-01-01
err0
PREAI
errTuan. A. Vu; Giang. H. Le; Canh. D. Dao; Lan. Q. Dang; Kien. T. Nguyen; Quang. K. Nguyen; Phuong. T. Dang; Hoa. T. K. Tran; Quang. T. Duong; Tuyen. V. Nguyen; Gun. D. Lee
err分享
err收藏
学者 查看更多内容