Online Deception Detection Refueled by Real World Data Collection
The lack of large realistic datasets presents a bottleneck in online deception detection studies. In this paper, we apply a data collection method based on social network analysis to quickly identify high-quality deceptive and truthful online reviews from Amazon. The dataset contains more than 10,000 deceptive reviews and is diverse in product domains and reviewers. Using this dataset, we explore effective general features for online deception detection that perform well across domains. We demonstrate that with generalized features - advertising speak and writing complexity scores - deception detection performance can be further improved by adding additional deceptive reviews from assorted domains in training. Finally, reviewer level evaluation gives an interesting insight into different deceptive reviewers' writing styles.
Code (0)
등록된 구현이 없습니다.
Tasks
Deception DetectionSimilar Papers 제목 키워드 기반
Can Deception Detection Go Deeper? Dataset, Evaluation, and Benchmark for Deception Reasoning
Deception detection has attracted increasing attention due to its importance in real-world scenarios. Its main goal is to detect deceptive behaviors from multimodal clues such as gestures, facial expressions, prosody, et…
Deception DetectionSentenceUnsupervised Audio-Visual Subspace Alignment for High-Stakes Deception Detection
Automated systems that detect deception in high-stakes situations can enhance societal well-being across medical, social work, and legal domains. Existing models for detecting high-stakes deception in videos have been su…
Deception DetectionTransfer LearningVocal Bursts Intensity PredictionHidden in Plain Sight: Evaluation of the Deception Detection Capabilities of LLMs in Multimodal Settings
Detecting deception in an increasingly digital world is both a critical and challenging task. In this study, we present a comprehensive evaluation of the automated deception detection capabilities of Large Language Model…
Deception DetectionXNote: Benchmarking Automated Community Notes Generation for Image-based Contextual Deception
Community Notes have emerged as an effective crowd-sourced mechanism for combating online deception on social media platforms. However, its reliance on human contributors limits both the timeliness and scalability. In th…
UNIDECOR: A Unified Deception Corpus for Cross-Corpus Deception Detection
Verbal deception has been studied in psychology, forensics, and computational linguistics for a variety of reasons, like understanding behaviour patterns, identifying false testimonies, and detecting deception in online …
Cross-corpusDeception DetectionDomain Generalization