arrow
Return

Mining top-k high average-utility itemsets based on breadth-first search

delete2023-10-27
delete0
PRE
AI
X
Xuan Liu
G
Genlang Chen
F
Fangyu Wu *
S
Shiting Wen
W
Wanli Zuo
DOI:10.1007/s10489-023-05076-4delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
High average-utility itemset mining is a subfield of data mining that has extensive practical applications. However, it is difficult for users to determine a proper minimum threshold because they cannot accurately predict the number of patterns mined at a given threshold. To address this issue, top-k high average-utility itemset mining has been proposed where k is the number of high average-utility itemsets to be mined. In this paper, we design an effective algorithm (named ETAUIM) for finding top-k high average-utility itemsets. ETAUIM employs a breadth-first search strategy to efficiently explore the search space, and it utilizes a tighter upper bound instead of the average-utility upper bound to limit the search space. Additionally, ETAUIM removes irrelevant items during the mining process and utilizes an early abandoning strategy to terminate unnecessary join operations in advance. To evaluate the proposed algorithm, extensive experiments were conducted on six sparse datasets and two dense datasets. Four state-of-the-art algorithms were used for comparison. The experimental results show that ETAUIM has excellent performance and scalability. Moreover, ETAUIM always performs better for sparse datasets.
Keywords:
Top-k high average-utility itemsets
Breadth-first search
High average-utility itemset
Data mining

Journal

Applied Intelligence cover
Applied Intelligence
IF:
3.5
Papers:
7.5K
Citations:
1.7W

Organization

N
Ningbotech University
Scholars:
1.0K
Papers: 777
Citations: 3
N
Ningbo University
Scholars:
2.6W
Papers: 1.8W
Citations: 2.4W