返回
High-confidence rule mining for Microarray analysis
DOI:10.1109/TCBB.2007.1050.png)
摘要
En 中文
We present an association rule mining method for mining high-confidence rules, which describe interesting gene relationships from microarray data sets. Microarray data sets typically contain an order of magnitude more genes than experiments, rendering many data mining methods impractical as they are optimized for sparse data sets. A new family of row-enumeration rule mining algorithms has emerged to facilitate mining in dense data sets. These algorithms rely on pruning infrequent relationships to reduce the search space by using the support measure. This major shortcoming results in the pruning of many potentially interesting rules with low support but high confidence. We propose a new row-enumeration rule mining method, MAXCONF, to mine high-confidence rules from microarray data. MAXCONF is a support-free algorithm that directly uses the confidence measure to effectively prune the search space. Experiments on three microarray data sets show that MAXCONF outperforms support-based rule mining with respect to scalability and rule extraction. Furthermore, detailed biological analyses demonstrate the effectiveness of our approach-the rules discovered by MAXCONF are substantially more interesting and meaningful compared with support-based methods.
Keyword:
data mining
association rules
high-confidence rule mining
microarray analysis
期刊
I
IF:
3.4
论文数:
3.3K
被引数:
6.4K
机构
暂无机构信息
引用论文
Comprehensive identification of cell cycle-regulated genes of the yeast Saccharomyces cerevisiae by microarray hybridization通过微阵列杂交全面鉴定酿酒酵母的细胞周期调控基因
The Gene Ontology (GO) database and informatics resource基因本体论 (GO) 数据库和信息学资源
NUCLEIC ACIDS RESEARCH
IF13.1

