paper-with-me

Papers

ActTraitBench: Quantifying the Knowledge-Decision Gap in Large Language Models via Human-Grounded Behavioral Validation

2026-05-28 · Yutong Yang, Chenxi Miao, Weikang Li, Yunfang Wu arxiv

While Large Language Models (LLMs) can convincingly simulate personas in explicit self-reports, they often deviate in implicit behavioral decisions, revealing a substantial Knowledge-Decision Gap ($G_{\mathrm{KD}}$). Existing benchmarks struggle to measure this discrepancy due to limited construct validity, multidimensional entanglement, and distributional biases in LLM-based evaluation. To address these issues, we propose ActTraitBench, a human-grounded evaluation framework for measuring personality consistency in LLMs. Grounded in empirical human data, ActTraitBench establishes one-to-one mappings between psychometric facets and behavioral paradigms and applies Distributional Calibration via Quantile Mapping to reduce distributional mismatch between LLM-judge scores and human responses. Experiments on 14 mainstream LLMs reveal substantial knowledge-decision gaps and show that assigned personas are reflected more consistently in self-reports than in behavioral decisions for most evaluated models. To mitigate this gap, we further introduce the Chain of Cognitive Alignment (CoCA), an inference-time intervention that reduces $G_{\mathrm{KD}}$ for 12 of the 13 models with paired results. Code and resources are available at https://github.com/Selina233/ActTraitBench.

📄 PDF Abstract BibTeX arXiv:2605.29791

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SafeDrive: Knowledge- and Data-Driven Risk-Sensitive Decision-Making for Autonomous Vehicles with Large Language Models

2024-12-17 · Zhiyuan Zhou, Heye Huang, Boqi Li, Shiyue Zhao 외

Recent advancements in autonomous vehicles (AVs) use Large Language Models (LLMs) to perform well in normal driving scenarios. However, ensuring safety in dynamic, high-risk environments and managing safety-critical long…

Autonomous DrivingAutonomous VehiclesDecision Making

Reinforcement Learning for Aligning Large Language Models Agents with Interactive Environments: Quantifying and Mitigating Prompt Overfitting

2024-10-25 · Mohamed Salim Aissi, Clement Romac, Thomas Carta, Sylvain Lamprier 외

Reinforcement learning (RL) is a promising approach for aligning large language models (LLMs) knowledge with sequential decision-making tasks. However, few studies have thoroughly investigated the impact on LLM agents ca…

Decision MakingReinforcement Learning (RL)SensitivitySequential Decision Making

Geometry of Decision Making in Language Models

2025-11-25 · Abhinav Joshi, Divyanshu Bhatt, Ashutosh Modi arxiv

Large Language Models (LLMs) show strong generalization across diverse tasks, yet the internal decision-making processes behind their predictions remain opaque. In this work, we study the geometry of hidden representatio…

Question AnsweringDecision Making

Scaling Laws for Task-Specific LLM Distillation

2026-06-23 · Lavinia Ghita, Dhruv Desai, Ioana Boier arxiv

Large Language Models (LLMs) achieve strong performance across a growing range of domains, yet their scale poses deployment challenges in applications where latency and cost constraints are critical. This paper derives e…

General Knowledge

Quantifying Cognitive Bias Induction in LLM-Generated Content

2025-07-03 · Abeer Alessa, Param Somane, Akshaya Lakshminarasimhan, Julian Skirzynski 외 arxiv

Large language models (LLMs) are integrated into applications like shopping reviews, summarization, or medical diagnosis support, where their use affects human decisions. We investigate the extent to which LLMs expose us…

Medical Diagnosis