paper-with-me

Papers

Assessing Large Language Models for Online Extremism Research: Identification, Explanation, and New Knowledge

2024-08-29 · Beidi Dong, Jin R. Lee, Ziwei Zhu, Balassubramanian Srinivasan

The United States has experienced a significant increase in violent extremism, prompting the need for automated tools to detect and limit the spread of extremist ideology online. This study evaluates the performance of Bidirectional Encoder Representations from Transformers (BERT) and Generative Pre-Trained Transformers (GPT) in detecting and classifying online domestic extremist posts. We collected social media posts containing "far-right" and "far-left" ideological keywords and manually labeled them as extremist or non-extremist. Extremist posts were further classified into one or more of five contributing elements of extremism based on a working definitional framework. The BERT model's performance was evaluated based on training data size and knowledge transfer between categories. We also compared the performance of GPT 3.5 and GPT 4 models using different prompts: na\"ive, layperson-definition, role-playing, and professional-definition. Results showed that the best performing GPT models outperformed the best performing BERT models, with more detailed prompts generally yielding better results. However, overly complex prompts may impair performance. Different versions of GPT have unique sensitives to what they consider extremist. GPT 3.5 performed better at classifying far-left extremist posts, while GPT 4 performed better at classifying far-right extremist posts. Large language models, represented by GPT models, hold significant potential for online extremism classification tasks, surpassing traditional BERT models in a zero-shot setting. Future research should explore human-computer interactions in optimizing GPT models for extremist detection and classification tasks to develop more efficient (e.g., quicker, less effort) and effective (e.g., fewer errors or mistakes) methods for identifying extremist content.

📄 PDF Abstract BibTeX arXiv:2408.16749

Code (0)

등록된 구현이 없습니다.

Tasks

Transfer Learning

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Cosine Annealing Cosine Annealing is a type of learning rate schedule that has the effect of starting with a large learning rate that is relatively rapidly decreased to a minimum value before…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Adam 설명 없음
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Weight Decay 설명 없음
Attention 설명 없음
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…

Similar Papers 제목 키워드 기반

Ideological Orientation and Extremism Detection in Online Social Networking Sites: A Systematic Review

2024-10-28 · Intelligent Systems with Applications 2024 10 · Kamalakkannan Ravi, Jiann-Shiun Yuan

The rise of social networking sites has reshaped digital interactions, becoming fertile grounds for extremist ideologies, notably in the United States. Despite previous research, understanding and tackling online ideolog…

ArticlesLiterature MiningSystematic Literature Review

Unifying the Extremes: Developing a Unified Model for Detecting and Predicting Extremist Traits and Radicalization

2025-01-08 · Allison Lahnala, Vasudha Varadarajan, Lucie Flek, H. Andrew Schwartz 외

The proliferation of ideological movements into extremist factions via social media has become a global concern. While radicalization has been studied extensively within the context of specific ideologies, our ability to…

ThreatGram 101 - Extreme Telegram Replies Data with Threat Levels

2024-09-30 · Information Management and Big Data. SIMBig 2024. Communications in Computer and Information Science. Springer, Cham. 2024 9 · Kamalakkannan Ravi, Jiann-Shiun Yuan

With the growth of social media, threats in comments targeting public officials, entities, or organizations have become increasingly common. Previous research on threat detection has typically focused on broad categories…

Abusive LanguageHate Speech DetectionInformation RetrievalMisinformation+4

Quantifying Extreme Opinions on Reddit Amidst the 2023 Israeli-Palestinian Conflict

2024-12-14 · Alessio Guerra, Marcello Lepre, Oktay Karakus

This study investigates the dynamics of extreme opinions on social media during the 2023 Israeli-Palestinian conflict, utilising a comprehensive dataset of over 450,000 posts from four Reddit subreddits (r/Palestine, r/J…

A Weakly Supervised Classifier and Dataset of White Supremacist Language

2023-06-27 · Michael Miller Yoder, Ahmad Diab, David West Brown, Kathleen M. Carley

We present a dataset and classifier for detecting the language of white supremacist extremism, a growing issue in online hate speech. Our weakly supervised classifier is trained on large datasets of text from explicitly …