paper-with-me

홈 › Papers

Exploiting Auxiliary Data for Offensive Language Detection with Bidirectional Transformers

2021-08-01 · ACL (WOAH) 2021 8 · Sumer Singh, Sheng Li

Offensive language detection (OLD) has received increasing attention due to its societal impact. Recent work shows that bidirectional transformer based methods obtain impressive performance on OLD. However, such methods usually rely on large-scale well-labeled OLD datasets for model training. To address the issue of data/label scarcity in OLD, in this paper, we propose a simple yet effective domain adaptation approach to train bidirectional transformers. Our approach introduces domain adaptation (DA) training procedures to ALBERT, such that it can effectively exploit auxiliary data from source domains to improve the OLD performance in a target domain. Experimental results on benchmark datasets show that our approach, ALBERT (DA), obtains the state-of-the-art performance in most cases. Particularly, our approach significantly benefits underrepresented and under-performing classes, with a significant improvement over ALBERT.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Domain Adaptation

Similar Papers 제목 키워드 기반

GruPaTo at SemEval-2020 Task 12: Retraining mBERT on Social Media and Fine-tuned Offensive Language Models

2020-12-01 · SEMEVAL 2020 · Davide Colla, Tommaso Caselli, Valerio Basile, Jelena Mitrovi{\'c} 외

We introduce an approach to multilingual Offensive Language Detection based on the mBERT transformer model. We download extra training data from Twitter in English, Danish, and Turkish, and use it to re-train the model. …

Transfer Learning

One to rule them all: Towards Joint Indic Language Hate Speech Detection

2021-09-28 · Mehar Bhatia, Tenzin Singhay Bhotia, Akshat Agarwal, Prakash Ramesh 외

This paper is a contribution to the Hate Speech and Offensive Content Identification in Indo-European Languages (HASOC) 2021 shared task. Social media today is a hotbed of toxic and hateful conversations, in various lang…

AllHate Speech Detection

Hate Speech and Offensive Language Detection using an Emotion-aware Shared Encoder

2023-02-17 · Khouloud Mnassri, Praboda Rajapaksha, Reza Farahbakhsh, Noel Crespi

The rise of emergence of social media platforms has fundamentally altered how people communicate, and among the results of these developments is an increase in online use of abusive content. Therefore, automatically dete…

Hate Speech Detection

Cross-Cultural Transfer Learning for Chinese Offensive Language Detection

2023-03-31 · Li Zhou, Laura Cabello, Yong Cao, Daniel Hershcovich

Detecting offensive language is a challenging task. Generalizing across different cultures and languages becomes even more challenging: besides lexical, syntactic and semantic differences, pragmatic aspects such as cultu…

Cultural Vocal Bursts Intensity PredictionFew-Shot LearningTransfer Learning

COLD: A Benchmark for Chinese Offensive Language Detection

2022-01-16 · Jiawen Deng, Jingyan Zhou, Hao Sun, Chujie Zheng 외

Offensive language detection is increasingly crucial for maintaining a civilized social media platform and deploying pre-trained language models. However, this task in Chinese is still under exploration due to the scarci…