paper-with-me

홈 › Papers

How Well Do LLMs Predict Human Behavior? A Measure of their Pretrained Knowledge

2026-01-18 · Wayne Gao, Sukjin Han, Annie Liang arxiv

Large language models (LLMs) are increasingly used to predict human behavior. We propose a measure for evaluating how much knowledge a pretrained LLM brings to such a prediction: its equivalent sample size, defined as the amount of task-specific data needed to match the predictive accuracy of the LLM. We estimate this measure by comparing the prediction error of a fixed LLM in a given domain to that of flexible machine learning models trained on increasing samples of domain-specific data. We further provide a statistical inference procedure by developing a new asymptotic theory for cross-validated prediction error. Finally, we apply this method to the Panel Study of Income Dynamics. We find that LLMs encode considerable predictive information for some economic variables but much less for others, suggesting that their value as substitutes for domain-specific data differs markedly across settings.

📄 PDF Abstract BibTeX arXiv:2601.12343

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

To model human linguistic prediction, make LLMs less superhuman

2025-10-01 · Byung-Doh Oh, Tal Linzen arxiv

When we read, we make predictions about upcoming words; these predictions influence our reading behavior. The success of large language models (LLMs), which, like humans, make predictions about upcoming words, has motiva…

Predicting Effects, Missing Distributions: Evaluating LLMs as Human Behavior Simulators in Operations Management

2025-09-30 · Runze Zhang, Xiaowei Zhang, Mingyang Zhao arxiv

Large language models (LLMs) are increasingly used to simulate human behavior in business, economics, and the social sciences, offering a low-cost complement to laboratory experiments, field studies, and surveys. This pa…

Psychometric Predictive Power of Large Language Models

2023-11-13 · Tatsuki Kuribayashi, Yohei Oseki, Timothy Baldwin

Instruction tuning aligns the response of large language models (LLMs) with human preferences. Despite such efforts in human--LLM alignment, we find that instruction tuning does not always make LLMs human-like from a cog…

Behavior Alignment: A New Perspective of Evaluating LLM-based Conversational Recommender Systems

2024-04-17 · Dayu Yang, Fumian Chen, Hui Fang

Large Language Models (LLMs) have demonstrated great potential in Conversational Recommender Systems (CRS). However, the application of LLMs to CRS has exposed a notable discrepancy in behavior between LLM-based CRS and …

Conversational RecommendationRecommendation Systems

Probing Large Language Models from A Human Behavioral Perspective

2023-10-08 · Xintong Wang, Xiaoyu Li, Xingshan Li, Chris Biemann

Large Language Models (LLMs) have emerged as dominant foundational models in modern NLP. However, the understanding of their prediction processes and internal mechanisms, such as feed-forward networks (FFN) and multi-hea…

MemorizationProbing Language Models