paper-with-me

Papers

A Study on Bias Detection and Classification in Natural Language Processing

2024-08-14 · Ana Sofia Evans, Helena Moniz, Luísa Coheur

Human biases have been shown to influence the performance of models and algorithms in various fields, including Natural Language Processing. While the study of this phenomenon is garnering focus in recent years, the available resources are still relatively scarce, often focusing on different forms or manifestations of biases. The aim of our work is twofold: 1) gather publicly-available datasets and determine how to better combine them to effectively train models in the task of hate speech detection and classification; 2) analyse the main issues with these datasets, such as scarcity, skewed resources, and reliance on non-persistent data. We discuss these issues in tandem with the development of our experiments, in which we show that the combinations of different datasets greatly impact the models' performance.

📄 PDF Abstract BibTeX arXiv:2408.07479

Code (0)

등록된 구현이 없습니다.

Tasks

Bias DetectionHate Speech Detection

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

GUS-Net: Social Bias Classification in Text with Generalizations, Unfairness, and Stereotypes

2024-10-10 · Maximus Powers, Umang Mavani, Harshitha Reddy Jonala, Ansh Tiwari 외

The detection of bias in natural language processing (NLP) is a critical challenge, particularly with the increasing use of large language models (LLMs) in various domains. This paper introduces GUS-Net, an innovative ap…

Bias Detectiontoken-classificationToken Classification

Detecting Unintended Social Bias in Toxic Language Datasets

2022-10-21 · Nihar Sahoo, Himanshu Gupta, Pushpak Bhattacharyya

With the rise of online hate speech, automatic detection of Hate Speech, Offensive texts as a natural language processing task is getting popular. However, very little research has been done to detect unintended social b…

Detecting Unintended Social Bias in Toxic Language Datasets

2022-01-16 · ACL ARR January 2022 1 · Anonymous

Hate speech and offensive texts are examples of damaging online content that target or promote hatred towards a group or individual member based on their actual or perceived features of identification, such as race, reli…

From Detection to Mitigation: Addressing Gender Bias in Chinese Texts via Efficient Tuning and Voting-Based Rebalancing

2025-09-09 · Chengyan Wu, Yiqiang Cai, Yufei Cheng, Yun Xue arxiv

This paper presents our team's solution to Shared Task 7 of NLPCC-2025, which focuses on sentence-level gender bias detection and mitigation in Chinese. The task aims to promote fairness and controllability in natural la…

Bias Detection

Bridging Human and Model Perspectives: A Comparative Analysis of Political Bias Detection in News Media Using Large Language Models

2025-11-18 · Shreya Adrita Banik, Niaz Nafi Rahman, Tahsina Moiukh, Farig Sadeque arxiv

Detecting political bias in news media is a complex task that requires interpreting subtle linguistic and contextual cues. Although recent advances in Natural Language Processing (NLP) have enabled automatic bias classif…

Bias Detection