On Cross-Dataset Generalization in Automatic Detection of Online Abuse
NLP research has attained high performances in abusive language detection as a supervised classification task. While in research settings, training and test datasets are usually obtained from similar data samples, in practice systems are often applied on data that are different from the training set in topic and class distributions. Also, the ambiguity in class definitions inherited in this task aggravates the discrepancies between source and target datasets. We explore the topic bias and the task formulation bias in cross-dataset generalization. We show that the benign examples in the Wikipedia Detox dataset are biased towards platform-specific topics. We identify these examples using unsupervised topic modeling and manual inspection of topics' keywords. Removing these topics increases cross-dataset generalization, without reducing in-domain classification performance. For a robust dataset design, we suggest applying inexpensive unsupervised methods to inspect the collected data and downsize the non-generalizable content before manually annotating for class labels.
Code (0)
등록된 구현이 없습니다.
Tasks
Abusive Languagedomain classificationSimilar Papers 제목 키워드 기반
Improving Cross-Domain Hate Speech Generalizability with Emotion Knowledge
Reliable automatic hate speech (HS) detection systems must adapt to the in-flow of diverse new data to curtail hate speech. However, hate speech detection systems commonly lack generalizability in identifying hate speech…
Hate Speech DetectionHotter and Colder: A New Approach to Annotating Sentiment, Emotions, and Bias in Icelandic Blog Comments
This paper presents Hotter and Colder, a dataset designed to analyze various types of online behavior in Icelandic blog comments. Building on previous work, we used GPT-4o mini to annotate approximately 800,000 comments …
Sentiment AnalysisControlled Automatic Task-Specific Synthetic Data Generation for Hallucination Detection
We present a novel approach to automatically generate non-trivial task-specific synthetic datasets for hallucination detection. Our approach features a two-step generation-selection pipeline, using hallucination pattern …
HallucinationIn-Context LearningSynthetic Data GenerationGeneralizing Hate Speech Detection Using Multi-Task Learning: A Case Study of Political Public Figures
Automatic identification of hateful and abusive content is vital in combating the spread of harmful online content and its damaging effects. Most existing works evaluate models by examining the generalization error on tr…
Hate Speech DetectionMulti-Task LearningDialogID: A Dialogic Instruction Dataset for Improving Teaching Effectiveness in Online Environments
Online dialogic instructions are a set of pedagogical instructions used in real-world online educational contexts to motivate students, help understand learning materials, and build effective study habits. In spite of th…