Improving Cross-Domain Hate Speech Detection by Reducing the False Positive Rate
Hate speech detection is an actively growing field of research with a variety of recently proposed approaches that allowed to push the state-of-the-art results. One of the challenges of such automated approaches – namely recent deep learning models – is a risk of false positives (i.e., false accusations), which may lead to over-blocking or removal of harmless social media content in applications with little moderator intervention. We evaluate deep learning models both under in-domain and cross-domain hate speech detection conditions, and introduce an SVM approach that allows to significantly improve the state-of-the-art results when combined with the deep learning models through a simple majority-voting ensemble. The improvement is mainly due to a reduction of the false positive rate.
Code (0)
등록된 구현이 없습니다.
Tasks
BlockingDeep LearningHate Speech DetectionSimilar Papers 제목 키워드 기반
Improving Cross-Domain Hate Speech Generalizability with Emotion Knowledge
Reliable automatic hate speech (HS) detection systems must adapt to the in-flow of diverse new data to curtail hate speech. However, hate speech detection systems commonly lack generalizability in identifying hate speech…
Hate Speech DetectionExposing the limits of Zero-shot Cross-lingual Hate Speech Detection
Reducing and counter-acting hate speech on Social Media is a significant concern. Most of the proposed automatic methods are conducted exclusively on English and very few consistently labeled, non-English resources have …
Cross-Lingual TransferHate Speech DetectionTransfer LearningZero-Shot Cross-Lingual TransferLarge-Scale Hate Speech Detection with Cross-Domain Transfer
Hate speech towards people with different backgrounds is a major problem observed in social media. Although there are various attempts to detect hate speech automatically via supervised learning models, the performance o…
Hate Speech DetectionLarge-Scale Hate Speech Detection with Cross-Domain Transfer
The performance of hate speech detection models relies on the datasets on which the models are trained. Existing datasets are mostly prepared with a limited number of instances or hate domains that define hate topics. Th…
Hate Speech DetectionTransfer LearningMultilingual Auxiliary Tasks Training: Bridging the Gap between Languages for Zero-Shot Transfer of Hate Speech Detection Models
Zero-shot cross-lingual transfer learning has been shown to be highly challenging for tasks involving a lot of linguistic specificities or when a cultural gap is present between languages, such as in hate speech detectio…
Cross-Lingual TransferHate Speech Detectionnamed-entity-recognitionNamed Entity Recognition+4