PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
July 16, 2011332 citations

Learning to identify review spam

View Full Paper
FLFangtao LiMHMinlie HuangYYYi Yang

Key Points

Key points are not available for this paper at this time.

Abstract

In the past few years, sentiment analysis and opinion mining becomes a popular and important task. These studies all assume that their opinion resources are real and trustful. However, they may encounter the faked opinion or opinion spam problem. In this paper, we study this issue in the context of our product review mining system. On product review site, people may write faked reviews, called review spam, to promote their products, or defame their competitors’ products. It is important to identify and filter out the review spam. Previous work only focuses on some heuristic rules, such as helpfulness voting, or rating deviation, which limits the performance of this task. In this paper, we exploit machine learning methods to identify review spam. Toward the end, we manually build a spam collection from our crawled reviews. We first analyze the effect of various features in spam identification. We also observe that the review spammer consistently writes spam. This provides us another view to identify review spam: we can identify if the author of the review is spammer. Based on this observation, we provide a twoview semi-supervised method, co-training, to exploit the large amount of unlabeled data. The experiment results show that our proposed method is effective. Our designed machine learning methods achieve significant improvements in comparison to the heuristic baselines.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Li et al. (2011) studied this question.

synapsesocial.com/papers/6a156b2c814bf8ec9a4e902dhttps://doi.org/10.5591/978-1-57735-516-8/ijcai11-414
Ask AI
Helpful
Bookmark
Share
View Full Paper