paper-with-me

홈 › Papers

Exploring Data Augmentation Strategies for Hate Speech Detection in Roman Urdu

2022-06-01 · LREC 2022 6 · Ubaid Azam, Hammad Rizwan, Asim Karim

In an era where social media platform users are growing rapidly, there has been a marked increase in hateful content being generated; to combat this, automatic hate speech detection systems are a necessity. For this purpose, researchers have recently focused their efforts on developing datasets, however, the vast majority of them have been generated for the English language, with only a few available for low-resource languages such as Roman Urdu. Furthermore, what few are available have small number of samples that pertain to hateful classes and these lack variations in topics and content. Thus, deep learning models trained on such datasets perform poorly when deployed in the real world. To improve performance the option of collecting and annotating more data can be very costly and time consuming. Thus, data augmentation techniques need to be explored to exploit already available datasets to improve model generalizability. In this paper, we explore different data augmentation techniques for the improvement of hate speech detection in Roman Urdu. We evaluate these augmentation techniques on two datasets. We are able to improve performance in the primary metric of comparison (F1 and Macro F1) as well as in recall, which is impertinent for human-in-the-loop AI systems.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationHate Speech Detection

Similar Papers 제목 키워드 기반

Decoding Hate: Exploring Language Models' Reactions to Hate Speech

2024-10-01 · Paloma Piot, Javier Parapar

Hate speech is a harmful form of online expression, often manifesting as derogatory posts. It is a significant risk in digital environments. With the rise of Large Language Models (LLMs), there is concern about their pot…

Hate Speech Detection using Large Language Models with Data Augmentation and Feature Enhancement

2026-03-05 · Brian Jing Hong Nge, Stefan Su, Thanh Thi Nguyen, Campbell Wilson 외 arxiv

This paper evaluates data augmentation and feature enhancement techniques for hate speech detection, comparing traditional classifiers, e.g., Delta Term Frequency-Inverse Document Frequency (Delta TF-IDF), with transform…

Hate Speech DetectionData Augmentation

Enhancing social network hate detection using back translation and GPT-3 augmentations during training and test-time

2023-06-17 · Information Fusion 2023 6 · Seffi Cohen, Dan Presil, Or Katz, Ofir Arbili 외

Social media platforms have become an essential means of communication, but they also serve as a breeding ground for hateful content. Detecting hate speech accurately is challenging due to factors such as slang and impli…

Hate Speech Detection

Enhancing Social Network Hate Detection Using Back Translation and GPT-3 Augmentations During Training and Test-Time

2023-06-13 · Information Fusion 2023 6 · S Cohen, D Presil, O Katz, O Arbili 외

Social media platforms have become an essential means of communication, but they also serve as a breeding ground for hateful content. Detecting hate speech accurately is challenging due to factors such as slang and impli…

Hate Speech DetectionHate Speech Normalization

Trustworthy Hate Speech Detection Through Visual Augmentation

2024-09-20 · Ziyuan Yang, Ming Yan, Yingyu Chen, Hui Wang 외

The surge of hate speech on social media platforms poses a significant challenge, with hate speech detection~(HSD) becoming increasingly critical. Current HSD methods focus on enriching contextual information to enhance …

Hate Speech Detection