paper-with-me

홈 › Papers

Preference Reasoning under Indeterminacy in Large Language Models

2026-08-19 · Hadi Hosseini, Samarth Khanna, Xiyuan Wang arxiv

As large language models evolve into decision-making agents, the ability to reason over preferences becomes fundamental to alignment, coordination, and collective intelligence. Yet, unlike standard benchmarks, real-world preference reasoning is inherently indeterminate: information may be incomplete, and valid solutions may not exist. We argue that indeterminacy, rather than correctness alone, is a central challenge for AI reasoning. We formalize this challenge along two axes, (i) epistemic indeterminacy, arising from incomplete, partial, or expressive preferences, and (ii) structural indeterminacy, arising from the non-existence of solutions under standard social choice concepts. Across a hierarchy of tasks, we show that state-of-the-art language models systematically fail to distinguish between determined and undetermined instances, exhibiting miscalibrated reasoning even in verification settings.

📄 PDF Abstract BibTeX arXiv:2608.18631

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Fuzzy ROC

2019-03-04 · Giovanni Parmigiani

The fuzzy ROC extends Receiver Operating Curve (ROC) visualization to the situation where some data points, falling in an indeterminacy region, are not classified. It addresses two challenges: definition of sensitivity a…

SensitivitySpecificity

DetermLR: Augmenting LLM-based Logical Reasoning from Indeterminacy to Determinacy

2023-10-28 · Hongda Sun, Weikai Xu, Wei Liu, Jian Luan 외

Recent advances in large language models (LLMs) have revolutionized the landscape of reasoning tasks. To enhance the capabilities of LLMs to emulate human reasoning, prior studies have focused on modeling reasoning steps…

Logical Reasoning

A Framework for Evaluating LLMs Under Task Indeterminacy

2024-11-21 · Luke Guerdan, Hanna Wallach, Solon Barocas, Alexandra Chouldechova

Large language model (LLM) evaluations often assume there is a single correct response -- a gold label -- for each item in the evaluation corpus. However, some tasks can be ambiguous -- i.e., they provide insufficient in…

Language ModelingLanguage ModellingLarge Language Model

Visual Indeterminacy in GAN Art

2019-10-10 · Aaron Hertzmann

This paper explores visual indeterminacy as a description for artwork created with Generative Adversarial Networks (GANs). Visual indeterminacy describes images which appear to depict real scenes, but, on closer examinat…

Image Generation

CodePMP: Scalable Preference Model Pretraining for Large Language Model Reasoning

2024-10-03 · Huimu Yu, Xing Wu, Weidong Yin, Debing Zhang 외

Large language models (LLMs) have made significant progress in natural language understanding and generation, driven by scalable pretraining and advanced finetuning. However, enhancing reasoning abilities in LLMs, partic…

GSM8KLanguage ModelingLanguage ModellingLarge Language Model+5