返回
Feature selection for noisy variation patterns using kernel principal component analysis
DOI:10.1016/j.knosys.2014.08.027.png)
摘要
En 中文
Kernel Principal Component Analysis (KPCA) is a technique widely used to understand and visualize nonlinear variation patterns by inverse mapping the projected data from a high-dimensional feature space back to the original input space. Variation patterns often occur in a small number of relevant features out of the overall set of features that are recorded in the data. It is, therefore, crucial to discern this set of relevant features that define the pattern. Here we propose a feature selection procedure that augments KPCA to obtain importance estimates of the features given the noisy training data. Our feature selection strategy involves projecting the data points onto sparse random vectors for calculating the kernel matrix. We then match pairs of such projections, and determine the preimages of the data with and without a feature, thereby trying to identify the importance of that feature. Thus, preimages' differences within pairs are used to identify the relevant features. An advantage of our method is it can be used with any suitable KPCA algorithm. Moreover, the computations can be parallelized easily leading to significant speedup. We demonstrate our method on several simulated and real data sets, and compare the results to alternative approaches in the literature. (C) 2014 Elsevier B.V. All rights reserved.
Keyword:
Nonlinear PCA
Kernel feature space
Preimages
Variation patterns
Feature ensembles
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
K
IF:
7.6
论文数:
1.3W
被引数:
4.5W
机构
引用论文
High incidence of brain and other nervous system cancer identified in two mining counties, 2001–2015
Fault detection of batch processes using multiway kernel principal component analysis基于多路核主元分析的间歇过程故障检测

