paper-with-me

홈 › Papers

Generating the Modal Worker: A Cross-Model Audit of Race and Gender in LLM-Generated Personas Across 41 Occupations

2025-10-23 · Ilona van der Linden, Sahana Kumar, Arnav Dixit, Aadi Sudan, Smruthi Danda, David C. Anastasiu, Kai Lukoff arxiv

As generative AI tools are increasingly used to portray people in professional roles, understanding their racial and gender representational biases is critical. We audit over 1.5 million occupational personas generated by four major large language models (GPT-4, Gemini 2.5, DeepSeek V3.1, and Mistral-medium) across 41 U.S. occupations. Comparing these personas against U.S. Bureau of Labor Statistics (BLS) data, we find that models generate demographics with less variation than real-world data, functionally compressing each occupation toward a dominant demographic profile rather than representing population-level variation. A shift/exaggeration decomposition reveals the structure of these distortions: White (-31 percentage points) and Black (-9 pp) workers are consistently underrepresented, while Hispanic (+17 pp) and Asian (+12 pp) workers are overrepresented, with stereotype exaggeration amplifying existing occupational segregation. These distortions are often extreme, including near-total portrayals of housekeepers as Hispanic and the near-erasure of Black workers from many occupations. Because these patterns recur across models with different institutional and cultural origins, they suggest shared structural sources of bias rather than model-specific artifacts. We argue that auditing generative AI requires evaluation frameworks that examine how synthetic populations systematically reshape demographic visibility across social roles.

📄 PDF Abstract BibTeX arXiv:2510.21011

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

RW-Post: Auditable Evidence-Grounded Multimodal Fact-Checking in the Wild

2025-12-28 · Danni Xu, Shaojing Fan, Harry Cheng, Mohan Kankanhalli arxiv

Multimodal misinformation increasingly leverages visual persuasion, where repurposed or manipulated images strengthen misleading text. We introduce RW-Post, a post-aligned text--image benchmark for real-world multimodal …

Visual Grounding

The Silicon Ceiling: Auditing GPT's Race and Gender Biases in Hiring

2024-05-07 · Lena Armstrong, Abbey Liu, Stephen MacNeil, Danaë Metaxa

Large language models (LLMs) are increasingly being introduced in workplace settings, with the goals of improving efficiency and fairness. However, concerns have arisen regarding these models' potential to reflect or exa…

Fairness

RW-Post: Auditable Evidence-Grounded Multimodal Fact-Checking in the Wild

2026-05-11 · Danni Xu, Shaojing Fan, Harry Cheng, Mohan Kankanhalli arxiv

Multimodal misinformation increasingly leverages visual persuasion, where repurposed or manipulated images strengthen misleading text. We introduce \textbf{RW-Post}, a post-aligned \textbf{text--image benchmark} for real…

Visual Grounding

Detecting Dataset Bias in Medical AI: A Generalized and Modality-Agnostic Auditing Framework

2025-03-13 · Nathan Drenkow, Mitchell Pavlak, Keith Harrigian, Ayah Zirikly 외

Data-driven AI is establishing itself at the center of evidence-based medicine. However, reports of shortcomings and unexpected behavior are growing due to AI's reliance on association-based learning. A major reason for …

AttributeLesion ClassificationMortality PredictionSkin Lesion Classification

TraceAV-Bench: Benchmarking Multi-Hop Trajectory Reasoning over Long Audio-Visual Videos

2026-05-08 · Hengyi Feng, Hao Liang, Mingrui Chen, Bohan Zeng 외 arxiv

Real-world audio-visual understanding requires chaining evidence that is sparse, temporally dispersed, and split across the visual and auditory streams, whereas existing benchmarks largely fail to evaluate this capabilit…

Multimodal Reasoning