Return
Information entropy based sample reduction for support vector data description
DOI:10.1016/j.asoc.2018.02.053.png)
Abstract
En 中文
Support vector data description (SVDD) is one of the most attractive methods in one-class classification (OCC), especially in solving problems in novelty detection. SVDD helps to deal with the classification witha large amount of target data and few outlier data. However, the huge computational complexity in kernel mapping makes it hard to be applied in use, as the number of target data increases. In order to reduce the size of the training data samples, we introduce a method called information entropy based sample reduction for support vector data description (IESRSVDD). In this method, the information entropy is calculated for the distribution of each data sample. The distance between each two samples is utilized to evaluate the probability of uncertainty for each sample. The samples with higher entropy values are considered to be near the boundary of the data distribution in kernel space, and likely to become support vectors. All samples with their entropy values lower than a threshold are excluded. An updated objective function of conventional SVDD is used in this method for sample reduction. The innovative highlights of the proposed IESRSVDD are: (i) reducing the training samples based on information entropy, (ii) introducing the sample reduction to SVDD in order to speed up the training process, and (iii) having the feasibility and effectiveness of IESRSVDD validated and analyzed. The experiment results show the proposed method can achieve a faster training speed by reducing the scale of the training set. The computing time is significantly reduced by 50-75% and the accuracy in classification is improved. (C) 2018 Elsevier B.V. All rights reserved.
Keywords:
Support vector data description
Information entropy
Sample reduction
One-class classification
AI Summary
Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.
Journal
IF:
6.6
Papers:
1.4W
Citations:
4.8W
Organization
No organization information available
Cited Papers
A boundary method for outlier detection based on support vector domain description
PATTERN RECOGNITION
IF7.6

