返回
Simulated annealing for supervised gene selection
DOI:10.1007/s00500-010-0597-8.png)
摘要
En 中文
Genomic data, and more generally biomedical data, are often characterized by high dimensionality. An input selection procedure can attain the two objectives of highlighting the relevant variables (genes) and possibly improving classification results. In this paper, we propose a wrapper approach to gene selection in classification of gene expression data using simulated annealing along with supervised classification. The proposed approach can perform global combinatorial searches through the space of all possible input subsets, can handle cases with numerical, categorical or mixed inputs, and is able to find (sub-)optimal subsets of inputs giving low classification errors. The method has been tested on publicly available bioinformatics data sets using support vector machines and on a mixed type data set using classification trees. We also propose some heuristics able to speed up the convergence. The experimental results highlight the ability of the method to select minimal sets of relevant features.
Keyword:
Input selection
DNA microarrays
Gene selection
Support vector machines
Classification trees
Simulated annealing
期刊
IF:
2.5
论文数:
1.0W
被引数:
2.1W

