paper-with-me

홈 › Papers

Manufactured Divisiveness: Decomposing the Hostile Content of Seven Social Media Influence Operations

2026-07-16 · Emilio Ferrara arxiv

State-backed influence operations are routinely measured as high-prevalence sources of `hate'' and toxicity.'' We argue those rates rest on a measurement error: the detectors behind them are validated to catch a broader definition inclusive of hostility or divisiveness aimed at an out-group, and so over-attribute hate to content better described as partisan or geopolitical invective. Across 25.08M tweets from seven government-attributed campaigns in the Twitter Information Operations archive (8,275 accounts), we separate hate from the other forms of divisiveness. We first validate a two-prompt LLM-based detector, matching human labels at Cohen's $κ=0.82$, to identify the broader hostility; we then develop an auditable rule, agreeing with an expert at $κ=0.52$, to further classify this content (5,457 posts) into three sub-categories. About 50.1% are identity-based attacks on people, whereas 30.4% are partisan attacks and 19.5% invective against states and their foreign policy. Reporting all of it as hate therefore overstates hate roughly twofold; only 18.7% is both identity-based and dehumanizing or inciting. Six of seven campaigns sort into three regimes that a single `hate'' rate flattens, namely identity hate (RU-op and IRA, both Russia-attributed), geopolitical invective (both Iran operations), and partisan divisiveness (both Venezuela operations). We call the shared product $manufactured divisiveness$. The line to separate these constructs itself remains unsettled: on the hardest cases three independent human experts agree only moderately (pairwise $κ=0.37$--$0.50$), and the best of nineteen LLM models tops out at $κ=0.601$ against the experts' majority. Our findings can help redefine the study of hate in the context of influence campaigns and broader online discourse.

📄 PDF Abstract BibTeX arXiv:2607.14491

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Measuring and Controlling Divisiveness in Rank Aggregation

2023-06-14 · Rachael Colley, Umberto Grandi, César Hidalgo, Mariana Macedo 외

In rank aggregation, members of a population rank issues to decide which are collectively preferred. We focus instead on identifying divisive issues that express disagreements among the preferences of individuals. We ana…

Decision Making

Hostility Detection in Hindi leveraging Pre-Trained Language Models

2021-01-14 · Ojasv Kamal, Adarsh Kumar, Tejas Vaidhya

Hostile content on social platforms is ever increasing. This has led to the need for proper detection of hostile posts so that appropriate action can be taken to tackle them. Though a lot of work has been done recently i…

Fake News DetectionHate Speech DetectionTransfer Learning

Interpreting Deep Neural Networks with Relative Sectional Propagation by Analyzing Comparative Gradients and Hostile Activations

2020-12-07 · Woo-Jeoung Nam, Jaesik Choi, Seong-Whan Lee

The clear transparency of Deep Neural Networks (DNNs) is hampered by complex internal structures and nonlinear transformations along deep hierarchies. In this paper, we propose a new attribution method, Relative Sectiona…

Walk in Wild: An Ensemble Approach for Hostility Detection in Hindi Posts

2021-01-15 · Chander Shekhar, Bhavya Bagla, Kaushal Kumar Maurya, Maunendra Sankar Desarkar

As the reach of the internet increases, pejorative terms started flooding over social media platforms. This leads to the necessity of identifying hostile content on social media platforms. Identification of hostile conte…

Binary ClassificationClassificationGeneral ClassificationMulti-class Classification

Evaluation of Deep Learning Models for Hostility Detection in Hindi Text

2021-01-11 · Ramchandra Joshi, Rushabh Karnavat, Kaustubh Jirapure, Raviraj Joshi

The social media platform is a convenient medium to express personal thoughts and share useful information. It is fast, concise, and has the ability to reach millions. It is an effective place to archive thoughts, share …

Multi-Label ClassificationMUlTI-LABEL-ClASSIFICATIONText DetectionWord Embeddings