paper-with-me

홈 › Papers

Artificial Intolerance: Stigmatizing Language in Clinical Documentation Skews Large Language Model Decision-Making

2026-05-17 · Jen-tse Huang, Didi Zhou, Faith Kamau, Amy Oh, Anne R. Links, Mark Dredze, Mary Catherine Beach, Somnath Saha arxiv

Large Language Models (LLMs) are increasingly deployed in high-stakes domains such as clinical decision support and medical documentation. However, the robustness of these models against subtle linguistic variations, specifically stigmatizing language (SL) commonly found in human-authored clinical notes, remains critically under-explored. In this work, we investigate whether frontier LLMs inherit and propagate this human bias when processing clinical text. We systematically evaluate nine frontier LLMs across four stigmatized medical conditions, utilizing clinical vignettes injected with varying intensities and phenotypes of SL (doubt, blame, and maligning). Our results demonstrate that all evaluated models exhibit substantial bias, with clinical decision-making significantly skewed towards less aggressive patient management. Notably, we observe a high sensitivity to linguistic framing, where a single SL sentence is sufficient to alter model outputs, revealing a clear dose-response relationship. Furthermore, we evaluate standard prompt-based mitigation strategies, including Chain-of-Thought (CoT) reasoning and model self-debiasing. These approaches show limited efficacy; models struggle to explicitly identify SL while remaining implicitly influenced by it. Our findings expose a critical vulnerability in current LLMs regarding fairness and robustness in clinical NLP, underscoring the need for rigorous algorithmic guardrails to prevent the automation of health disparities.

📄 PDF Abstract BibTeX arXiv:2605.17228

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Understanding Stigmatizing Language in Clinical Documentation: A Paired Comparison of Ambient AI Drafts and Clinician Finalized Notes

2026-04-14 · Yiliang Zhou, Yawen Guo, Sairam Sutari, Jasmine Dhillon 외 arxiv

Ambient artificial intelligence (AI) documentation tools are increasingly deployed to reduce clinician documentation burden, but their implications for biased language in clinical notes remain unclear. We conducted a lar…

Fine-Tune, Don't Prompt, Your Language Model to Identify Biased Language in Clinical Notes

2026-02-16 · Isotta Landi, Eugenia Alleva, Nicole Bussola, Rebecca M. Cohen 외 arxiv

Clinical documentation can contain emotionally charged language with stigmatizing or privileging valences. We present a framework for detecting and classifying such language as stigmatizing, privileging, or neutral. We c…

Prompt EngineeringBias Detection

Understanding Stigmatizing Language Lexicons: A Comparative Analysis in Clinical Contexts

2025-09-09 · Yiliang Zhou, Di Hu, Tianchu Lyu, Jasmine Dhillon 외 arxiv

Stigmatizing language results in healthcare inequities, yet there is no universally accepted or standardized lexicon defining which words, terms, or phrases constitute stigmatizing language in healthcare. We conducted a …

Semantic SimilaritySentiment Analysis

CARE-SD: Classifier-based analysis for recognizing and eliminating stigmatizing and doubt marker labels in electronic health records: model development and validation

2024-05-08 · Drew Walker, Annie Thorne, Sudeshna Das, Jennifer Love 외

Objective: To detect and classify features of stigmatizing and biased language in intensive care electronic health records (EHRs) using natural language processing techniques. Materials and Methods: We first created a le…

Intelligent Clinical Documentation: Harnessing Generative AI for Patient-Centric Clinical Note Generation

2024-05-28 · Anjanava Biswas, Wrick Talukdar

Comprehensive clinical documentation is crucial for effective healthcare delivery, yet it poses a significant burden on healthcare professionals, leading to burnout, increased medical errors, and compromised patient safe…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition