arrow
返回

Scalable parallel data mining for association rules

delete2000-01-01
delete94
PRE
AI
E
Eui-Hong Han *
G
George Karypis
K
Kumar, V
DOI:10.1109/69.846289delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In this paper, we propose two new parallel formulations of the Apriori algorithm that is used for computing association rules. These new formulations, IDD and HD, address the shortcomings of two previously proposed parallel formulations CD and DD. Unlike the CD algorithm, the IDD algorithm partitions the candidate set intelligently among processors to efficiently parallelize the step of building the hash tree. The IDD algorithm also eliminates the redundant work inherent in DD, and requires substantially smaller communication overhead than DD. But IDD suffers from the added cost due to communication of transactions among processors. HD is a hybrid algorithm that combines the advantages of Co and DD. Experimental results on a 128-processor Gray T3E show that HD scales just as well as the CD algorithm with respect to the number of transactions, and scales as well as IDD with respect to increasing candidate set size.
Keyword:
data mining
parallel processing
association rules
load balance
scalability
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Knowledge and Data Engineering 封面图
IEEE Transactions on Knowledge and Data Engineering
IF:
10.4
论文数:
6.8K
被引数:
3.2W

机构

暂无机构信息
引用论文

引用论文

Semi-analytic contact technique in a non-linear parametric model order reduction method for gear simulations
err2017-06-17
err0
PREAI
errNiccolò Cappellini; Tommaso Tamarozzi; Bart Blockmans; Jakob Fiszer; Francesco Cosco; Wim Desmet
err分享
err收藏
err分享
err收藏
err分享
err收藏
没有更多内容