paper-with-me

홈 › Papers

On Cross-Dataset Generalization in Automatic Detection of Online Abuse

2020-10-14 · EMNLP (ALW) 2020 11 · Isar Nejadgholi, Svetlana Kiritchenko

NLP research has attained high performances in abusive language detection as a supervised classification task. While in research settings, training and test datasets are usually obtained from similar data samples, in practice systems are often applied on data that are different from the training set in topic and class distributions. Also, the ambiguity in class definitions inherited in this task aggravates the discrepancies between source and target datasets. We explore the topic bias and the task formulation bias in cross-dataset generalization. We show that the benign examples in the Wikipedia Detox dataset are biased towards platform-specific topics. We identify these examples using unsupervised topic modeling and manual inspection of topics' keywords. Removing these topics increases cross-dataset generalization, without reducing in-domain classification performance. For a robust dataset design, we suggest applying inexpensive unsupervised methods to inspect the collected data and downsize the non-generalizable content before manually annotating for class labels.

📄 PDF Abstract BibTeX arXiv:2010.07414

Code (0)

등록된 구현이 없습니다.

Tasks

Abusive Languagedomain classification

Similar Papers 제목 키워드 기반

Improving Cross-Domain Hate Speech Generalizability with Emotion Knowledge

2023-11-24 · Shi Yin Hong, Susan Gauch

Reliable automatic hate speech (HS) detection systems must adapt to the in-flow of diverse new data to curtail hate speech. However, hate speech detection systems commonly lack generalizability in identifying hate speech…

Hate Speech Detection

Hotter and Colder: A New Approach to Annotating Sentiment, Emotions, and Bias in Icelandic Blog Comments

2025-02-24 · Steinunn Rut Friðriksdóttir, Dan Saattrup Nielsen, Hafsteinn Einarsson

This paper presents Hotter and Colder, a dataset designed to analyze various types of online behavior in Icelandic blog comments. Building on previous work, we used GPT-4o mini to annotate approximately 800,000 comments …

Sentiment Analysis

Controlled Automatic Task-Specific Synthetic Data Generation for Hallucination Detection

2024-10-16 · Yong Xie, Karan Aggarwal, Aitzaz Ahmad, Stephen Lau

We present a novel approach to automatically generate non-trivial task-specific synthetic datasets for hallucination detection. Our approach features a two-step generation-selection pipeline, using hallucination pattern …

HallucinationIn-Context LearningSynthetic Data Generation

Generalizing Hate Speech Detection Using Multi-Task Learning: A Case Study of Political Public Figures

2022-08-22 · Lanqin Yuan, Marian-Andrei Rizoiu

Automatic identification of hateful and abusive content is vital in combating the spread of harmful online content and its damaging effects. Most existing works evaluate models by examining the generalization error on tr…

Hate Speech DetectionMulti-Task Learning

DialogID: A Dialogic Instruction Dataset for Improving Teaching Effectiveness in Online Environments

2022-06-24 · Jiahao Chen, Shuyan Huang, Zitao Liu, Weiqi Luo

Online dialogic instructions are a set of pedagogical instructions used in real-world online educational contexts to motivate students, help understand learning materials, and build effective study habits. In spite of th…