paper-with-me

홈 › Papers

Detecting Anti-Semitic Hate Speech using Transformer-based Large Language Models

2024-05-06 · Dengyi Liu, Minghao Wang, Andrew G. Catlin

Academic researchers and social media entities grappling with the identification of hate speech face significant challenges, primarily due to the vast scale of data and the dynamic nature of hate speech. Given the ethical and practical limitations of large predictive models like ChatGPT in directly addressing such sensitive issues, our research has explored alternative advanced transformer-based and generative AI technologies since 2019. Specifically, we developed a new data labeling technique and established a proof of concept targeting anti-Semitic hate speech, utilizing a variety of transformer models such as BERT (arXiv:1810.04805), DistillBERT (arXiv:1910.01108), RoBERTa (arXiv:1907.11692), and LLaMA-2 (arXiv:2307.09288), complemented by the LoRA fine-tuning approach (arXiv:2106.09685). This paper delineates and evaluates the comparative efficacy of these cutting-edge methods in tackling the intricacies of hate speech detection, highlighting the need for responsible and carefully managed AI applications within sensitive contexts.

📄 PDF Abstract BibTeX arXiv:2405.03794

Code (0)

등록된 구현이 없습니다.

Tasks

Hate Speech Detection

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Attention 설명 없음
Weight Decay 설명 없음
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
Residual Connection 설명 없음
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
WordPiece 설명 없음

Similar Papers 제목 키워드 기반

Using LLMs to discover emerging coded antisemitic hate-speech in extremist social media

2024-01-19 · Dhanush Kikkisetti, Raza Ul Mustafa, Wendy Melillo, Roberto Corizzo 외

Online hate speech proliferation has created a difficult problem for social media platforms. A particular challenge relates to the use of coded language by groups interested in both creating a sense of belonging for its …

Language ModelingLanguage ModellingLarge Language ModelSemantic Similarity+1

A Group-Specific Approach to NLP for Hate Speech Detection

2023-04-21 · Karina Halevy

Automatic hate speech detection is an important yet complex task, requiring knowledge of common sense, stereotypes of protected groups, and histories of discrimination, each of which may constantly evolve. In this paper,…

Common Sense ReasoningEthicsHate Speech Detection

Antisemitic Messages? A Guide to High-Quality Annotation and a Labeled Dataset of Tweets

2023-04-28 · Gunther Jikeli, Sameer Karali, Daniel Miehling, Katharina Soemer

One of the major challenges in automatic hate speech detection is the lack of datasets that cover a wide range of biased and unbiased messages and that are consistently labeled. We propose a labeling procedure that addre…

Hate Speech Detection

Evaluating Large Language Models for Antisemitic Incident Classification

2026-07-06 · Karina Halevy, Julia Mendelsohn, Chan Young Park, Yulia Tsvetkov 외 arxiv

Addressing hate and violence in society requires timely detection of hateful events from public reporting, but automated identification of hateful events remains underexplored. We introduce the task of hateful event dete…

The Role of Context in Detecting the Target of Hate Speech

2022-10-01 · TRAC (COLING) 2022 10 · Ilia Markov, Walter Daelemans

Online hate speech detection is an inherently challenging task that has recently received much attention from the natural language processing community. Despite a substantial increase in performance, considerable challen…

Hate Speech DetectionLanguage ModelingLanguage Modelling