paper-with-me

홈 › Papers

Disentangling Dialect from Social Bias via Multitask Learning to Improve Fairness

2024-06-14 · Maximilian Spliethöver, Sai Nikhil Menon, Henning Wachsmuth

Dialects introduce syntactic and lexical variations in language that occur in regional or social groups. Most NLP methods are not sensitive to such variations. This may lead to unfair behavior of the methods, conveying negative bias towards dialect speakers. While previous work has studied dialect-related fairness for aspects like hate speech, other aspects of biased language, such as lewdness, remain fully unexplored. To fill this gap, we investigate performance disparities between dialects in the detection of five aspects of biased language and how to mitigate them. To alleviate bias, we present a multitask learning approach that models dialect language as an auxiliary task to incorporate syntactic and lexical variations. In our experiments with African-American English dialect, we provide empirical evidence that complementing common learning approaches with dialect modeling improves their fairness. Furthermore, the results suggest that multitask learning achieves state-of-the-art performance and helps to detect properties of biased language more reliably.

📄 PDF Abstract BibTeX arXiv:2406.09977

Code (1)

webis-de/acl-24

Tasks

Fairness

Similar Papers 제목 키워드 기반

Adversarial Multitask Learning for Joint Multi-Feature and Multi-Dialect Morphological Modeling

2019-10-28 · ACL 2019 7 · Nasser Zalmout, Nizar Habash

Morphological tagging is challenging for morphologically rich languages due to the large target space and the need for more training data to minimize model sparsity. Dialectal variants of morphologically rich languages s…

Morphological TaggingTransfer Learning

Dialect Diversity in Text Summarization on Twitter

2020-07-15 · Vijay Keswani, L. Elisa Celis

Discussions on Twitter involve participation from different communities with different dialects and it is often necessary to summarize a large number of posts into a representative sample to provide a synopsis. Yet, any …

AttributeDiversityExtractive SummarizationLanguage Identification+1

Side-by-side Comparison Amplifies Dialect Bias in Language Models

2026-05-23 · Kritee Kondapally, Claire J. Smerdon, Pooja C. Patel, Ogheneyoma Akoni 외 arxiv

Language models (LMs) can exhibit biases based on variations in their dialects, even in the absence of a dialect label, a behavior known as covert dialect bias. In this work, we quantify covert dialect bias in online dis…

Decision Making

Automatic Speech Recognition Biases in Newcastle English: an Error Analysis

2025-06-19 · Dana Serditova, Kevin Tang, Jochen Steffens

Automatic Speech Recognition (ASR) systems struggle with regional dialects due to biased training which favours mainstream varieties. While previous research has identified racial, age, and gender biases in ASR, regional…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Diversityspeech-recognition+1

Mitigating Biases in Toxic Language Detection through Invariant Rationalization

2021-06-14 · ACL (WOAH) 2021 8 · Yung-Sung Chuang, Mingye Gao, Hongyin Luo, James Glass 외

Automatic detection of toxic language plays an essential role in protecting social media users, especially minority groups, from verbal abuse. However, biases toward some attributes, including gender, race, and dialect, …

Natural Language Understanding