paper-with-me

Papers

MALCOM: Generating Malicious Comments to Attack Neural Fake News Detection Models

2020-09-01 · Thai Le, Suhang Wang, Dongwon Lee

In recent years, the proliferation of so-called "fake news" has caused much disruptions in society and weakened the news ecosystem. Therefore, to mitigate such problems, researchers have developed state-of-the-art models to auto-detect fake news on social media using sophisticated data science and machine learning techniques. In this work, then, we ask "what if adversaries attempt to attack such detection models?" and investigate related issues by (i) proposing a novel threat model against fake news detectors, in which adversaries can post malicious comments toward news articles to mislead fake news detectors, and (ii) developing MALCOM, an end-to-end adversarial comment generation framework to achieve such an attack. Through a comprehensive evaluation, we demonstrate that about 94% and 93.5% of the time on average MALCOM can successfully mislead five of the latest neural detection models to always output targeted real and fake news labels. Furthermore, MALCOM can also fool black box fake news detectors to always output real news labels 90% of the time on average. We also compare our attack model with four baselines across two real-world datasets, not only on attack performance but also on generated quality, coherency, transferability, and robustness.

📄 PDF Abstract BibTeX arXiv:2009.01048

Code (1)

lethaiq/MALCOM 공식 구현 pytorch

Tasks

ArticlesComment GenerationFake News Detection

Similar Papers 제목 키워드 기반

Group-Adaptive Adversarial Learning for Robust Fake News Detection Against Malicious Comments

2025-10-10 · Zhao Tong, Chunlin Gong, Yimeng Gu, Haichao Shi 외 arxiv

Online fake news profoundly distorts public judgment and erodes trust in social platforms. While existing detectors achieve competitive performance on benchmark datasets, they remain notably vulnerable to malicious comme…

Fake News Detection

A Robust Opinion Spam Detection Method Against Malicious Attackers in Social Media

2020-08-19 · Amir Jalaly Bidgolya, Zoleikha Rahmaniana

Online reviews are potent sources for industry owners and buyers, however opportunistic people may try to destruct or promote their desired product by publishing fake comments named spam opinion. So far, many models have…

Spam detection

Attack Graph Convolutional Networks by Adding Fake Nodes

2018-10-25 · ICLR 2019 5 · Xiaoyun Wang, Minhao Cheng, Joe Eaton, Cho-Jui Hsieh 외

In this paper, we study the robustness of graph convolutional networks (GCNs). Previous work have shown that GCNs are vulnerable to adversarial perturbation on adjacency or feature matrices of existing nodes; however, su…

Shilling Recommender Systems by Generating Side-feature-aware Fake User Profiles

2025-09-22 · Yuanrong Wang, Yingpeng Du arxiv

Recommender systems (RS) greatly influence users' consumption decisions, making them attractive targets for malicious shilling attacks that inject fake user profiles to manipulate recommendations. Existing shilling metho…

Turn Fake into Real: Adversarial Head Turn Attacks Against Deepfake Detection

2023-09-03 · Weijie Wang, Zhengyu Zhao, Nicu Sebe, Bruno Lepri

Malicious use of deepfakes leads to serious public concerns and reduces people's trust in digital media. Although effective deepfake detectors have been proposed, they are substantially vulnerable to adversarial attacks.…

DeepFake DetectionFace Swapping