paper-with-me

홈 › Papers

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes

2026-04-24 · Rubén Garzón, Pauline Baron, Vincent Grari, Jonne Kamphorst, Michael Bernstein, Marcin Detyniecki arxiv

Large language models (LLM) agents may offer tools to predict human responses to surveys. A common technique for defining these agents uses only demographics, for example country, age, gender, employment status, income, education and marital status. We compare the predictive accuracy of demographic agents to that of survey agents defined with a larger set of in-domain survey responses. We test both approaches in predicting responses to the multidisciplinary, cross-national Survey of Health, Ageing and Retirement in Europe (SHARE), focusing on five variables from three policy-relevant constructs around personal finance. In these three constructs, we observe that, compared to survey agents trained on broader data, demographics-only agents (1) exhibited a central tendency bias, skewing answers toward population means, and (2) were unrealistically accurate, failing to reproduce the incorrect answers and "don't know" responses typical of human respondents. These performance differences are further substantiated through the replication of a hierarchical regression analysis from prior retirement planning research. Agents based solely on demographic information reproduce the outcome that financial risk tolerance, future time perspective, and knowledge of retirement planning each are predictive of retirement savings. However, only the survey-anchored agents succeed in reproducing the interaction among these three factors. These findings suggest caution in using only demographics to define LLM agents for predicting survey responses.

📄 PDF Abstract BibTeX arXiv:2605.16303

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Accurate and Data-Efficient Toxicity Prediction when Annotators Disagree

2024-10-16 · Harbani Jaggi, Kashyap Murali, Eve Fleisig, Erdem Biyik

When annotators disagree, predicting the labels given by individual annotators can capture nuances overlooked by traditional label aggregation. We introduce three approaches to predicting individual annotator ratings on …

Collaborative FilteringIn-Context LearningPredictionSurvey

Aligning Language Models to User Opinions

2023-05-24 · EunJeong Hwang, Bodhisattwa Prasad Majumder, Niket Tandon

An important aspect of developing LLMs that interact with humans is to align models' behavior to their users. It is possible to prompt an LLM into behaving as a certain persona, especially a user group or ideological per…

Open-Ended Question Answering

Questioning the Survey Responses of Large Language Models

2023-06-13 · Ricardo Dominguez-Olmedo, Moritz Hardt, Celestine Mendler-Dünner

Surveys have recently gained popularity as a tool to study large language models. By comparing survey responses of models to those of human reference populations, researchers aim to infer the demographics, political opin…

Multiple-choiceSurvey

Predicting Demographics of High-Resolution Geographies with Geotagged Tweets

2017-01-22 · Omar Montasser, Daniel Kifer

In this paper, we consider the problem of predicting demographics of geographic units given geotagged Tweets that are composed within these units. Traditional survey methods that offer demographics estimates are usually …

SurveyVocal Bursts Intensity Prediction

CoMPosT: Characterizing and Evaluating Caricature in LLM Simulations

2023-10-17 · Myra Cheng, Tiziano Piccardi, Diyi Yang

Recent work has aimed to capture nuances of human behavior by using LLMs to simulate responses from particular demographics in settings like social science experiments and public opinion surveys. However, there are curre…

Caricature