paper-with-me

Papers

When Agents Say One Thing and Do Another: Validating Elicited Beliefs from LLMs

2026-02-06 · Khurram Yamin, Jingjing Tang, Santiago Cortes-Gomez, Amit Sharma, Eric Horvitz, Bryan Wilder arxiv

Large language models (LLMs) are increasingly deployed in high-stakes settings where good decisions require forming beliefs over the probability of unknown outcomes. However, it is unclear whether LLMs act as if they hold coherent beliefs when making decisions, or if so, how we could validate models' reports of such beliefs. We propose a decision-theoretic framework that elicits both probability judgments and decisions from an agent and tests their mutual consistency. Formally, our methods characterize whether it is possible for the actions to be produced by a ``near-rational" decision maker who holds the elicited probability as their true belief. We show that, perhaps surprisingly, this formalization implies empirically testable conditions even without any assumption about the agent's utility function. Applying our framework to stylized clinical diagnosis tasks, we find that models' reported beliefs are demonstrably imperfect summaries of the information revealed in their decisions, but that the discrepancies are small for the strongest models.

📄 PDF Abstract BibTeX arXiv:2602.06286

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Modeling User Empathy Elicited by a Robot Storyteller

2021-07-29 · Leena Mathur, Micol Spitale, Hao Xi, Jieyun Li 외

Virtual and robotic agents capable of perceiving human empathy have the potential to participate in engaging and meaningful human-machine interactions that support human well-being. Prior research in computational empath…

Understanding the Cognitive Complexity in Language Elicited by Product Images

2024-09-25 · Yan-Ying Chen, Shabnam Hakimi, Monica Van, Francine Chen 외

Product images (e.g., a phone) can be used to elicit a diverse set of consumer-reported features expressed through language, including surface-level perceptual attributes (e.g., "white") and more complex ones, like perce…

Descriptive

Uncertainty Aware Learning for Language Model Alignment

2024-06-07 · Yikun Wang, Rui Zheng, Liang Ding, Qi Zhang 외

As instruction-tuned large language models (LLMs) evolve, aligning pretrained foundation models presents increasing challenges. Existing alignment strategies, which typically leverage diverse and high-quality data source…

GSM8KLanguage ModelingLanguage Modellingmodel

e-Genia3 An AgentSpeak extension for empathic agents

2022-08-01 · Joaquin Taverner, Emilio Vivancos, Vicente Botti

In this paper, we present e-Genia3 an extension of AgentSpeak to provide support to the development of empathic agents. The new extension modifies the agent's reasoning processes to select plans according to the analyzed…

Network Classifiers With Output Smoothing

2019-10-30 · Elsa Rizk, Roula Nassif, Ali H. Sayed

This work introduces two strategies for training network classifiers with heterogeneous agents. One strategy promotes global smoothing over the graph and a second strategy promotes local smoothing over neighbourhoods. It…