arrow
Return

Boundary sample extraction support vector data description: a novel anomaly detection method for large-scale data

delete2025-01-21
delete0
PRE
AI
王小飞 cover
王小飞 (Xiaofei Wang)
Y
Yongzhan Chen *
F
Fenglei Xu
高彦丽 (Yanli Gao)
W
Wang, Yuanxin
Y
Yuchuan Qiao
王强 (Qiang Wang)
DOI:10.1088/1361-6501/ada6eadelete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Support vector data description (SVDD) has been effectively used in many anomaly detection problems. Equipped with kernel functions, its training complexity grows exponentially with the increase in training data, which makes it less practical for large-scale datasets. In this paper, we propose a boundary sample extraction SVDD (BSE-SVDD) anomaly detection method based on data reduction, aiming for use with large-scale data. Firstly, the BSE mechanism is established on the basis of demonstrating that the spatial position of the sample is related to its corresponding Lagrange multiplier. Then, the BSE mechanism is used to search for the local optimal solution of the Lagrange multiplier, and all the training samples are sorted. Finally, the top p of ranked samples are extracted as boundary samples for training, while most of the training samples that may be non-support vectors are removed. Compared with SVDD studies based on data reduction, the experimental results on large-scale datasets show that BSE-SVDD can obtain comparable classification accuracy with greatly improved training speed.
Keywords:
support vector data description
boundary samples extraction
data reduction
stochastic gradient descent
large-scale data

Journal

Measurement Science and Technology cover
Measurement Science and Technology
IF:
3.4
Papers:
2.6K
Citations:
2.3W

Organization

N
naval aeronaut univ
Scholars:
10
Papers: 4
Citations: 0