paper-with-me

Papers

Ambiguity Collapse by LLMs: A Taxonomy of Epistemic Risks

2026-03-06 · Shira Gur-Arieh, Angelina Wang, Sina Fazelpour arxiv

Large language models (LLMs) are increasingly used to make sense of ambiguous, open-textured, value-laden terms. Platforms routinely rely on LLMs for content moderation, asking them to label text based on disputed concepts like "hate speech" or "incitement"; hiring managers may use LLMs to rank who counts as "qualified"; and AI labs increasingly train models to self-regulate under constitutional-style ambiguous principles such as "biased" or "legitimate". This paper introduces ambiguity collapse: a phenomenon that occurs when an LLM encounters a term that genuinely admits multiple legitimate interpretations, yet produces a singular resolution, in ways that bypass the human practices through which meaning is ordinarily negotiated, contested, and justified. Drawing on interdisciplinary accounts of ambiguity as a productive epistemic resource, we develop a taxonomy of the epistemic risks posed by ambiguity collapse at three levels: process (foreclosing opportunities to deliberate, develop cognitive skills, and shape contested terms), output (distorting the concepts and reasons agents act upon), and ecosystem (reshaping shared vocabularies, interpretive norms, and how concepts evolve over time). We illustrate these risks through three case studies, and conclude by sketching multi-layer mitigation principles spanning training, institutional deployment design, interface affordances, and the management of underspecified prompts, with the goal of designing systems that surface, preserve, and responsibly govern ambiguity.

📄 PDF Abstract BibTeX arXiv:2603.05801

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CLAMBER: A Benchmark of Identifying and Clarifying Ambiguous Information Needs in Large Language Models

2024-05-20 · Tong Zhang, Peixin Qin, Yang Deng, Chen Huang 외

Large language models (LLMs) are increasingly used to meet user information needs, but their effectiveness in dealing with user queries that contain various types of ambiguity remains unknown, ultimately risking user tru…

Epistemic Diversity and Knowledge Collapse in Large Language Models

2025-10-05 · Dustin Wright, Sarah Masud, Jared Moore, Srishti Yadav 외 arxiv

Large language models (LLMs) tend to generate homogenous texts, which may impact the diversity of knowledge generated across different outputs. Given their potential to replace existing forms of knowledge acquisition, th…

A taxonomy of epistemic injustice in the context of AI and the case for generative hermeneutical erasure

2025-04-10 · Warmhold Jan Thomas Mollema

Whether related to machine learning models' epistemic opacity, algorithmic classification systems' discriminatory automation of testimonial prejudice, the distortion of human beliefs via the 'hallucinations' of generativ…

Philosophy

Knowledge Collapse in LLMs: When Fluency Survives but Facts Fail under Recursive Synthetic Training

2025-09-05 · Figarri Keisha, Zekun Wu, Ze Wang, Adriano Koshiyama 외 arxiv

Large language models increasingly rely on synthetic data due to human-written content scarcity, yet recursive training on model-generated outputs leads to model collapse, a degenerative process threatening factual relia…

Computational Efficiency

Epistemic Uncertainty Is Not the Reducible Kind

2026-06-10 · Robin Young arxiv

The standard taxonomy of predictive uncertainty defines epistemic uncertainty as the part removable by collecting more data, while the standard measure identifies it with a mutual-information term. We prove the definitio…