paper-with-me

홈 › Papers

Impacts of Racial Bias in Historical Training Data for News AI

2025-12-18 · Rahul Bhargava, Malene Hornstrup Jespersen, Emily Boardman Ndulue, Vivica Dsouza arxiv

AI technologies have rapidly moved into business and research applications that involve large text corpora, including computational journalism research and newsroom settings. These models, trained on extant data from various sources, can be conceptualized as historical artifacts that encode decades-old attitudes and stereotypes. This paper investigates one such example trained on the broadly-used New York Times Annotated Corpus to create a multi-label classifier. Our use in research settings surfaced the concerning "blacks" thematic topic label. Through quantitative and qualitative means we investigate this label's use in the training corpus, what concepts it might be encoding in the trained classifier, and how those concepts impact our model use. Via the application of explainable AI methods, we find that the "blacks" label operates partially as a general "racism detector" across some minoritized groups. However, it performs poorly against expectations on modern examples such as COVID-19 era anti-Asian hate stories, and reporting on the Black Lives Matter movement. This case study of interrogating embedded biases in a model reveals how similar applications in newsroom settings can lead to unexpected outputs that could impact a wide variety of potential uses of any large language model-story discovery, audience targeting, summarization, etc. The fundamental tension this exposes for newsrooms is how to adopt AI-enabled workflow tools while reducing the risk of reproducing historical biases in news coverage.

📄 PDF Abstract BibTeX arXiv:2512.16901

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Studying Bias in GANs through the Lens of Race

2022-09-06 · Vongani H. Maluleke, Neerja Thakkar, Tim Brooks, Ethan Weber 외

In this work, we study how the performance and evaluation of generative image models are impacted by the racial composition of their training datasets. By examining and controlling the racial distributions in various tra…

Detecting Racial Bias in Jury Selection

2021-03-22 · Jack Dunn, Ying Daisy Zhuo

To support the 2019 U.S. Supreme Court case "Flowers v. Mississippi", APM Reports collated historical court records to assess whether the State exhibited a racial bias in striking potential jurors. This analysis used bac…

feature selection

Gender and Racial Stereotype Detection in Legal Opinion Word Embeddings

2022-03-24 · Sean Matthews, John Hudzina, Dawn Sepehr

Studies have shown that some Natural Language Processing (NLP) systems encode and replicate harmful biases with potential adverse ethical effects in our society. In this article, we propose an approach for identifying ge…

Question AnsweringWord Embeddings

A Race Bias Free Face Aging Model for Reliable Kinship Verification

2025-09-18 · Ali Nazari, Bardiya Kariminia, Mohsen Ebrahimi Moghaddam arxiv

The age gap in kinship verification addresses the time difference between the photos of the parent and the child. Moreover, their same-age photos are often unavailable, and face aging models are racially biased, which im…

Kinship Verification

AI & Racial Equity: Understanding Sentiment Analysis Artificial Intelligence, Data Security, and Systemic Theory in Criminal Justice Systems

2022-01-03 · Alia Abbas

Various forms of implications of artificial intelligence that either exacerbate or decrease racial systemic injustice have been explored in this applied research endeavor. Taking each thematic area of identifying, analyz…

Decision MakingSentiment Analysis