paper-with-me

홈 › Papers

Predicting Effects, Missing Distributions: Evaluating LLMs as Human Behavior Simulators in Operations Management

2025-09-30 · Runze Zhang, Xiaowei Zhang, Mingyang Zhao arxiv

Large language models (LLMs) are increasingly used to simulate human behavior in business, economics, and the social sciences, offering a low-cost complement to laboratory experiments, field studies, and surveys. This paper evaluates how well LLMs replicate human behavior in operations management. Using nine published behavioral-operations experiments, we assess LLM performance along two dimensions: whether LLM-generated data reproduce the original hypothesis-test outcomes, and whether their full response distributions align with human data, measured by Wasserstein distance. We find that LLMs often replicate hypothesis-level effects, suggesting that they can capture salient decision biases and behavioral regularities. However, their response distributions frequently diverge from human data, even for strong proprietary models, with dispersion mismatch playing an important role. We also examine two lightweight mitigation strategies: chain-of-thought prompting and hyperparameter tuning. Both can reduce distributional misalignment, and appropriate tuning can sometimes allow smaller or open-source models to match or outperform larger proprietary systems.

📄 PDF Abstract BibTeX arXiv:2510.03310

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Evaluating Interventional Reasoning Capabilities of Large Language Models

2024-04-08 · Tejas Kasetty, Divyat Mahajan, Gintare Karolina Dziugaite, Alexandre Drouin 외

Numerous decision-making tasks require estimating causal effects under interventions on different parts of a system. As practitioners consider using large language models (LLMs) to automate decisions, studying their caus…

Causal InferenceDecision Making

A Quantitative Evaluation Framework for Missing Value Imputation Algorithms

2013-11-10 · Vinod Nair, Rahul Kidambi, Sundararajan Sellamanickam, S. Sathiya Keerthi 외

We consider the problem of quantitatively evaluating missing value imputation algorithms. Given a dataset with missing values and a choice of several imputation algorithms to fill them in, there is currently no principle…

ImputationMissing Values

Federated Causal Inference from Observational Data

2023-08-24 · Thanh Vinh Vo, Young Lee, Tze-Yun Leong

Decentralized data sources are prevalent in real-world applications, posing a formidable challenge for causal inference. These sources cannot be consolidated into a single entity owing to privacy constraints. The presenc…

Causal InferenceFederated LearningGaussian ProcessesMissing Values+1

LLM-based Online Prediction of Time-varying Graph Signals

2024-10-24 · Dayu Qin, Yi Yan, Ercan Engin Kuruoglu

In this paper, we propose a novel framework that leverages large language models (LLMs) for predicting missing values in time-varying graph signals by exploiting spatial and temporal smoothness. We leverage the power of …

Missing Values

Who Should I Engage with At What Time? A Missing Event Aware Temporal Graph Neural Network

2023-01-20 · Mingyi Liu, Zhiying Tu, Xiaofei Xu, Zhongjie Wang

Temporal graph neural network has recently received significant attention due to its wide application scenarios, such as bioinformatics, knowledge graphs, and social networks. There are some temporal graph neural network…

Graph Neural NetworkKnowledge GraphsLink PredictionPoint Processes