paper-with-me

Papers

LLM-Derived Preference Judgments Are Not Self-Consistent

2026-08-18 · Matthew T. Ford, Francis Bahk, Jingjing Wang, Adam S. Jovine, Tinghan Ye, David B. Shmoys, Peter I. Frazier arxiv

Agents increasingly interpret a person's natural-language preferences by querying an LLM for numerical preference judgments, e.g., by asking how much the person would be willing to pay for an item. A growing body of work estimates a utility function from these judgments and then chooses actions based on their estimated utility. This pipeline assumes the judgments are approximately self-consistent: that a single utility function can reproduce them. But are they? To study this question, we measure the self-consistency of cardinal LLM preference judgments. For example, the difference in stated willingness-to-pay between two items should match the stated payment that makes a person indifferent to exchanging them. We develop statistical tests and interpretable measures of how far observed responses depart from the best-fitting self-consistent utility function. Experiments with flight, apartment, and hotel examples across six LLMs reveal large persistent inconsistencies. This suggests that LLM-derived preference judgments cannot be faithfully summarized by a single utility function.

📄 PDF Abstract BibTeX arXiv:2608.17644

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Assessing top-$k$ preferences

2020-07-22 · Charles L. A. Clarke, Alexandra Vtyurina, Mark D. Smucker

Assessors make preference judgments faster and more consistently than graded judgments. Preference judgments can also recognize distinctions between items that appear equivalent under graded judgments. Unfortunately, pre…

Quantifying Label-Induced Bias in Large Language Model Self- and Cross-Evaluations

2025-08-28 · Muskan Saraf, Sajjad Rezvani Boroujeni, Justin Beaudry, Hossein Abedi 외 arxiv

Large language models (LLMs) are increasingly deployed as evaluators of text quality, yet the validity of their judgments remains underexplored. This study investigates systematic bias in self- and cross-model evaluation…

Beyond Scalar Reward Model: Learning Generative Judge from Preference Data

2024-10-01 · Ziyi Ye, Xiangsheng Li, Qiuchi Li, Qingyao Ai 외

Learning from preference feedback is a common practice for aligning large language models~(LLMs) with human value. Conventionally, preference data is learned and encoded into a scalar reward model that connects a value h…

DecipherPref: Analyzing Influential Factors in Human Preference Judgments via GPT-4

2023-05-24 · Yebowen Hu, Kaiqiang Song, Sangwoo Cho, Xiaoyang Wang 외

Human preference judgments are pivotal in guiding large language models (LLMs) to produce outputs that align with human values. Human evaluations are also used in summarization tasks to compare outputs from various syste…

Informativeness

Transitivity, Time Consumption, and Quality of Preference Judgments in Crowdsourcing

2021-04-18 · Kai Hui, Klaus Berberich

Preference judgments have been demonstrated as a better alternative to graded judgments to assess the relevance of documents relative to queries. Existing work has verified transitivity among preference judgments when co…