arrow
返回

Opinion spam detection: Using multi-iterative graph-based model

delete2020-01-01
delete58
PRE
AI
S
Shirin Noekhah *
N
Naomie Salim
DOI:10.1016/j.ipm.2019.102140delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The demand to detect opinionated spam, using opinion mining applications to prevent their damaging effects on e-commerce reputations is on the rise in many business sectors globally. The existing spam detection techniques in use nowadays, only consider one or two types of spam entities such as review, reviewer, group of reviewers, and product. Besides, they use a limited number of features related to behaviour, content and the relation of entities which reduces the detection's accuracy. Accordingly, these techniques mostly exploit synthetic datasets to analyse their model and are not able to be applied in the context of the real-world environment. As such, a novel graph-based model called Multi-iterative Graph-based opinion Spam Detection (MGSD) in which all various types of entities are considered simultaneously within a unified structure is proposed. Using this approach, the model reveals both implicit (i.e., similar entity's) and explicit (i.e., different entities') relationships. The MGSD model is able to evaluate the 'spamicity' effects of entities more efficiently given it applies a novel multi-iterative algorithm which considers different sets of factors to update the spamicity score of entities. To enhance the accuracy of the MGSD detection model, a higher number of existing weighted features along with the novel proposed features from different categories were selected using a combination of feature fusion techniques and machine learning (ML) algorithms. The MGSD model can also be generalised and applied in various opinionated documents due to employing domain independent features. The output of the MGSD model showed that our feature selection and feature fusion techniques showed a remarkable improvement in detecting spam. The findings of this study showed that MGSD could improve the accuracy of state-of-the-art ML and graph-based techniques by around 5.6% and 4.8%, respectively, also achieving an accuracy of 93% for the detection of spam detection in our synthetic crowdsourced dataset and 95.3% for Ott's crowdsourced dataset.
Keyword:
Opinion spam detection
Heterogeneous graph-based structure
Spammer
Group of spammers
Feature fusion
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

I
Information Processing and Management
IF:
6.9
论文数:
5.2K
被引数:
1.4W

机构

U
Universiti Teknologi Malaysia
学者数:
1.4W
论文数: 1.1W
被引数: 85
引用论文

引用论文

Biologically Inspired Soft Robot for Thumb Rehabilitation1
err2014-04-28
err0
PREAI
errPaxton Maeder-York; Tyler Clites; Emily Boggs; Ryan Neff; Panagiotis Polygerinos; Dónal Holland; Leia Stirling; Kevin Galloway; Catherine Wee; Conor Walsh
err分享
err收藏
Detecting positive and negative deceptive opinions using PU-learning使用PU学习检测积极和消极的欺骗性意见
err2015-07-01
err118
errOAAI
errHernandez Fusilier, Donato; Montes-y-Gomez, Manuel; Rosso, Paolo; Guzman Cabrera, Rafael
err分享
err收藏
Agency Problems and Risk Taking At Banks
err1997-01-01
err0
errOAAI
errRebecca S. Demsetz; Marc R. Saidenberg; Philip E. Strahan
err分享
err收藏
Shari'ah Compliance as a Matter for Financial Performance
err2019-12-27
err0
PREAI
errMd. Harun Ur Rashid; Md Hafij Ullah; Faruk Bhuiyan
err分享
err收藏
Coal Desulfurization with Acidithiobacillus ferrivorans, from Balya Acidic Mine Drainage
err2013-06-11
err0
PREAI
errPinar Aytar; Catherine M. Kay; Mehmet Burçin Mutlu; Ahmet Çabuk
err分享
err收藏
学者 查看更多内容