Online Deception Detection Refueled by Real World Data Collection

07/28/2017
by   Wenlin Yao, et al.
0

The lack of large realistic datasets presents a bottleneck in online deception detection studies. In this paper, we apply a data collection method based on social network analysis to quickly identify high-quality deceptive and truthful online reviews from Amazon. The dataset contains more than 10,000 deceptive reviews and is diverse in product domains and reviewers. Using this dataset, we explore effective general features for online deception detection that perform well across domains. We demonstrate that with generalized features - advertising speak and writing complexity scores - deception detection performance can be further improved by adding additional deceptive reviews from assorted domains in training. Finally, reviewer level evaluation gives an interesting insight into different deceptive reviewers' writing styles.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
05/31/2019

Spotting Collusive Behaviour of Online Fraud Groups in Customer Reviews

Online reviews play a crucial role in deciding the quality before purcha...
research
04/03/2020

Directions in Abusive Language Training Data: Garbage In, Garbage Out

Data-driven analysis and detection of abusive online content covers many...
research
05/31/2019

Spotting Collective Behaviour of Online Frauds in Customer Reviews

Online reviews play a crucial role in deciding the quality before purcha...
research
09/28/2020

Reactive Supervision: A New Method for Collecting Sarcasm Data

Sarcasm detection is an important task in affective computing, requiring...
research
12/31/2020

Learning from the Worst: Dynamically Generated Datasets to Improve Online Hate Detection

We present a first-of-its-kind large synthetic training dataset for onli...
research
07/03/2023

Internet of Things Fault Detection and Classification via Multitask Learning

This paper presents a comprehensive investigation into developing a faul...
research
12/06/2021

Letter-level Online Writer Identification

Writer identification (writer-id), an important field in biometrics, aim...

Please sign up or login with your details

Forgot password? Click here to reset