arrow
返回

Distributed mining of high utility time interval sequential patterns using mapreduce approach

delete2020-03-01
delete23
PRE
AI
S
Sumalatha Saleti *
R
R. B. V. Subramanyam
DOI:10.1016/j.eswa.2019.112967delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
High Utility Sequential Pattern mining (HUSP) algorithms aim to find all the high utility sequences from a sequence database. Due to the large explosion of data, recently few distributed algorithms have been designed for mining HUSPs based on the MapReduce framework. However, the existing HUSP algorithms such as USpan, HUS-Span and BigHUSP are able to predict only the order of items, they do not predict the time between the items, that is, they do not include the time intervals between the successive items. But in a real-world scenario, time interval patterns provide more valuable information than conventional high utility sequential patterns. Therefore, we propose a distributed high utility time interval sequential pattern mining (DHUTISP) algorithm using the MapReduce approach that is suitable for big data. DHUTISP creates a novel time interval utility linked list data structure (TIUL) to efficiently calculate the utility of the resulting patterns. Moreover, two utility upper bounds, namely, remaining utility upper bound (RUUB) and co-occurrence utility upper bound (CUUB) are proposed to prune the unpromising candidates. We conducted various experiments to prove the efficiency of the proposed algorithm over both the distributed and non-distributed approaches. The experimental results show the efficiency of DHUTISP over state-of-the-art algorithms, namely, BigHUSP, AHUS-P, PUSOM and UTMining_A. (C) 2019 Elsevier Ltd. All rights reserved.
Keyword:
Big data
High utility itemset mining
High utility sequential pattern mining
Time interval sequential pattern mining
Mapreduce framework
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Expert Systems with Applications 封面图
Expert Systems with Applications
IF:
7.5
论文数:
2.9W
被引数:
10.2W

机构

N
national institute of technology (nit system)
学者数:
4.0W
论文数: 3.7W
被引数: 31
引用论文

引用论文

A Taxonomy of Sequential Pattern Mining Algorithms
err2010-12-03
err229
PREAI
errMabroukeh, Nizar R.; Ezeife, C. I.
err分享
err收藏
err分享
err收藏
Evaluation of Diverse Convolutional Neural Networks and Training Strategies for Wheat Leaf Disease Identification with Field-Acquired Photographs
err2022-07-18
err0
errOAAI
errJiale Jiang; Haiyan Liu; Chen Zhao; Can He; Jifeng Ma; Tao Cheng; Yan Zhu; Weixing Cao; Xia Yao
err分享
err收藏
Mining summarization of high utility itemsets
err2015-08-01
err13
PREAI
errZhang, Xiong; Deng, Zhi-Hong
err分享
err收藏
An efficient algorithm to mine high average-utility itemsets
err2016-04-01
err71
PREAI
errLin, Jerry Chun-Wei; Li, Ting; Fournier-Viger, Philippe; Hong, Tzung-Pei; Zhan, Justin; Voznak, Miroslav
err分享
err收藏
err分享
err收藏
Applying the maximum utility measure in high utility sequential pattern mining
err2014-09-01
err92
PREAI
errLan, Guo-Cheng; Hong, Tzung-Pei; Tseng, Vincent S.; Wang, Shyue-Liang
err分享
err收藏
学者 查看更多内容