返回
Some issues on scalable feature selection
DOI:10.1016/S0957-4174(98)90049-5.png)
摘要
En 中文
Feature selection determines relevant features in the data. It is often applied in pattern classification, data mining, as well as machine learning. A special concern for feature selection nowadays is that the size of a database is normally very large, both vertically and horizontally. In addition, feature sets may grow as the data collection process continues. Effective solutions are needed to accommodate the practical demands. This paper concentrates on three issues: large number of features, large data size, and expanding feature set. For the first issue, we suggest a probabilistic algorithm to select features. For the second issue, we present a scalable probabilistic algorithm that expedites feature selection further and can scale up without sacrificing the quality of selected features. For the third issue, we propose an incremental algorithm that adapts to the newly extended feature set and captures ''concept drifts' by removing features from previously selected and newly added ones. We expect that research on scalable: feature selection will be extended to distributed and parallel computing and have impact on applications of data mining and machine learning. (C) 1998 Elsevier Science Ltd. All rights reserved.
Keyword:
features
probabilistic selection
scalability
large databases
pattern classification
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
7.5
论文数:
3.0W
被引数:
10.2W
机构
暂无机构信息
引用论文
A Study of Effects of Design Parameters on Transient Response and Injection Rate Shaping for a Common Rail Injector System设计参数对共轨喷油器系统瞬态响应和喷油率整形影响的研究
没有更多内容

