arrow
返回

Which standard classification algorithm has more stable performance for imbalanced network traffic data?

delete2023-10-26
delete0
PRE
AI
M
Ming Zheng *
K
Kai Ma
F
Fei Wang
X
Xiaowen Hu
俞庆英 (Qingying Yu)
L
Liangmin Guo
陈付龙 封面图
陈付龙 (Fulong Chen)
DOI:10.1007/s00500-023-09331-1delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Most standard classification algorithms are difficult to effectively learn and predict from imbalanced network traffic data, which usually leads to lower classification accuracy. To analyze the influence of imbalanced network traffic data on the performance of standard classification algorithms, the imbalanced data augmentation algorithms are first designed to obtain the imbalanced network traffic data set with gradually varying Imbalance Ratio (IR) and belonging to the same distribution. Then, to obtain more objective classification result and simplify the evaluation process, the evaluation metric AFG is used to evaluate the classification performance of standard classification algorithms based on area under the receiver operating characteristic curve (AUC), F-measure and G-mean. Finally, based on AFG and coefficient of variation (CV), performance stability of standard classification algorithms on imbalanced network traffic data is obtained. Experiments of eight widely used standard classification algorithms on 25 different imbalanced network traffic data demonstrate that the classification performance of GNB, RF and DT is unstable, while BNB, KNN, LR, GBDT, and SVC are relatively stable and not susceptible to imbalanced data. Especially, the KNN has the most stable classification performance. Also, the results are statistically confirmed by Friedman and Nemenyi post hoc statistical tests.
Keyword:
Imbalanced network traffic data
Data augmentation algorithms
Standard classification algorithms
Stable classification performance

期刊

Soft Computing 封面图
Soft Computing
IF:
2.5
论文数:
1.0W
被引数:
2.1W

机构

A
Anhui Normal University
学者数:
7.0K
论文数: 4.6K
被引数: 6.8K
引用论文

引用论文

Mineral composition of fruit by-products evaluated by neutron activation analysis
err2013-01-13
err0
PREAI
errGabriela de Matuoka e Chiocchetti; Elisabete A. De Nadai Fernandes; Márcio Arruda Bacchi; Rogério Augusto Pazim; Silvana Regina Vicino Sarriés; Thaís Melega Tomé
err分享
err收藏
The impact of class imbalance in classification performance metrics based on the binary confusion matrix
err2019-07-01
err590
errOAAI
errLuque, Amalia; Carrasco, Alejandro; Martin, Alejandro; de las Heras, Ana
err分享
err收藏
Contemporary evaluation of measurement uncertainties in vector network analysis
err2017-01-11
err0
PREAI
errMarkus Zeier; Johannes Hoffmann; Juerg Ruefenacht; Michael Wollensack
err分享
err收藏
Arsenic removal from aqueous solutions by adsorption using novel MIL-53(Fe) as a highly efficient adsorbent使用新型MIL-53(Fe) 作为高效吸附剂通过吸附从水溶液中去除砷
err2015-01-01
err0
PREAI
errTuan. A. Vu; Giang. H. Le; Canh. D. Dao; Lan. Q. Dang; Kien. T. Nguyen; Quang. K. Nguyen; Phuong. T. Dang; Hoa. T. K. Tran; Quang. T. Duong; Tuyen. V. Nguyen; Gun. D. Lee
err分享
err收藏
Preprocessed dynamic classifier ensemble selection for highly imbalanced drifted data streams
err2021-02-01
err68
PREAI
errZyblewski, Pawel; Sabourin, Robert; Wozniak, Michal
err分享
err收藏
Optimizing Weighted Extreme Learning Machines for imbalanced classification and application to credit card fraud detection
err2020-09-01
err98
PREAI
errZhu, Honghao; Liu, Guanjun; Zhou, Mengchu; Xie, Yu; Abusorrah, Abdullah; Kang, Qi
err分享
err收藏
err分享
err收藏
One-class support vector classifiers: A survey
err2020-05-01
err81
PREAI
errAlam, Shamshe; Sonbhadra, Sanjay Kumar; Agarwal, Sonali; Nagabhushan, P.
err分享
err收藏
学者 查看更多内容