Return
Data-dependent kernel machines for Microarray data classification
DOI:10.1109/TCBB.2007.1048.png)
Abstract
En 中文
One important application of gene expression analysis is to classify tissue samples according to their gene expression levels. Gene expression data are typically characterized by high dimensionality and small sample size, which makes the classification task quite challenging. In this paper, we present a data-dependent kernel for microarray data classification. This kernel function is engineered so that the class separability of the training data is maximized. A bootstrapping-based resampling scheme is introduced to reduce the possible training bias. The effectiveness of this adaptive kernel for microarray data classification is illustrated with a k-Nearest Neighbor (KNN) classifier. Our experimental study shows that the data-dependent kernel leads to a significant improvement in the accuracy of KNN classifiers. Furthermore, this kernel-based KNN scheme has been demonstrated to be competitive to, if not better than, more sophisticated classifiers such as Support Vector Machines (SVMs) and the Uncorrelated Linear Discriminant Analysis (ULDA) for classifying gene expression data.
Keywords:
microarray data analysis
cancer classification
kernel machines
kernel optimization
bootstrapping resampling
AI Summary
Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.
Journal
I
IF:
3.4
Papers:
3.3K
Citations:
6.4K
Organization
No organization information available
Cited Papers
A computational approach to identify genes for functional RNAs in genomic sequences
NUCLEIC ACIDS RESEARCH
IF13.1
Exploring a fiscal food policy: the case of diet and ischaemic heart disease Commentary: Alternative nutrition outcomes using a fiscal food policy
BMJ
IF0
Diffuse large B-cell lymphoma outcome prediction by gene-expression profiling and supervised machine learning
NATURE MEDICINE
IF50

