arrow
返回

Efficient high utility itemset mining using buffered utility-lists

delete2017-09-15
delete80
PRE
AI
Q
Quang-Huy Duong
P
Philippe Fournier‐Viger *
H
Heri Ramampiaro *
K
Kjetil Nørvåg
T
Thu-Lan Dam
DOI:10.1007/s10489-017-1057-2delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Discovering high utility itemsets in transaction databases is a key task for studying the behavior of customers. It consists of finding groups of items bought together that yield a high profit. Several algorithms have been proposed to mine high utility itemsets using various approaches and more or less complex data structures. Among existing algorithms, one-phase algorithms employing the utility-list structure have shown to be the most efficient. In recent years, the simplicity of the utility-list structure has led to the development of numerous utility-list based algorithms for various tasks related to utility mining. However, a major limitation of utility-list based algorithms is that creating and maintaining utility-lists are time consuming and can consume a huge amount of memory. The reasons are that numerous utility lists are built and that the utility-list intersection/join operation to construct a utility-list is costly. This paper addresses this issue by proposing an improved utility-list structure called utility-list buffer to reduce the memory consumption and speed up the join operation. This structure is integrated into a novel algorithm named ULB-Miner (Utility-List Buffer for high utility itemset Miner), which introduces several new ideas to more efficiently discover high utility itemsets. ULB-Miner uses the designed utility-list buffer structure to efficiently store and retrieve utility-lists, and reuse memory during the mining process. Moreover, the paper also introduces a linear time method for constructing utility-list segments in a utility-list buffer. An extensive experimental study on various datasets shows that the proposed algorithm relying on the novel utility-list buffer structure is highly efficient in terms of both execution time and memory consumption. The ULB-Miner algorithm is up to 10 times faster than the FHM and HUI-Miner algorithms and consumes up to 6 times less memory. Moreover, it performs well on both dense and sparse datasets.
Keyword:
Pattern mining
Itemset mining
Utility mining
Utility list
Utility list buffer
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Applied Intelligence 封面图
Applied Intelligence
IF:
3.5
论文数:
7.6K
被引数:
1.7W

机构

H
harbin institute of technology
学者数:
8.0W
论文数: 6.6W
被引数: 66
引用论文

引用论文

err分享
err收藏
Thromboelastography Predictive of Death in Trauma Patients
err2015-02-23
err0
errOAAI
errIan Kane; Alvin Ong; Fabio R Orozco; Zachary D Post; Luke S Austin; Kris E Radcliff
err分享
err收藏
Efficient Algorithms for Mining Top-K High Utility Itemsets
err2016-01-01
err167
PREAI
errTseng, Vincent S.; Wu, Cheng-Wei; Fournier-Viger, Philippe; Yu, Philip S.
err分享
err收藏
err分享
err收藏
Hydrogenation of prostaglandin unsaturated ketones over Ru-containing *BEA zeolites
err2004-01-01
err0
PREAI
errS.N. Coman; D.C. Radu; V.I. Parvulescu; Z. Sobalik; D.E. De Vos; P.A. Jacobs
err分享
err收藏
学者 查看更多内容