paper-with-me

홈 › Papers

Large Language Models as Virtual Survey Respondents: Evaluating Sociodemographic Response Generation

2025-09-08 · Jianpeng Zhao, Chenyu Yuan, Weiming Luo, Haoling Xie, Guangwei Zhang, Steven Jige Quan, Zixuan Yuan, Pengyang Wang, Denghui Zhang arxiv

Questionnaire-based surveys are foundational to social science research and public policymaking, yet traditional survey methods remain costly, time-consuming, and often limited in scale. Although prior work has explored large language models (LLMs) as virtual survey respondents, existing studies often address narrow task settings, focus on single sociological domains, or lack a unified evaluation framework that enables systematic comparison across diverse datasets and models. To address these gaps, we introduce two complementary task abstractions: Partial Attribute Simulation (PAS), where LLMs predict missing attributes from incomplete respondent profiles, and Full Attribute Simulation (FAS), where LLMs generate complete synthetic datasets under zero-context and context-enhanced conditions. Both are framed as diagnostic and exploratory tools rather than replacements for human data collection. We curate LLM-S^3 (Large Language Model-based Sociodemographic Survey Simulation), a benchmark spanning 11 real-world public datasets across four sociological domains, and evaluate GPT-3.5/4 Turbo and LLaMA 3.0/3.1-8B under zero-shot and few-shot settings. Our evaluation reveals consistent performance trends across model families, highlights failure modes in structured output generation, and demonstrates how context and prompt design affect simulation fidelity. Our code and dataset are available at: https://github.com/dart-lab-research/LLM-S-Cube-Benchmark

📄 PDF Abstract BibTeX arXiv:2509.06337

Code (0)

등록된 구현이 없습니다.

Tasks

Response Generation

Similar Papers 제목 키워드 기반

Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators

2025-07-08 · Sungjib Lim, Woojung Song, Eun-Ju Lee, Yohan Jo arxiv

As psychometric surveys are increasingly used to assess the traits of large language models (LLMs), the need for scalable survey item generation suited for LLMs has also grown. A critical challenge here is ensuring the c…

What Do Dialect Speakers Want? A Survey of Attitudes Towards Language Technology for German Dialects

2024-02-19 · Verena Blaschke, Christoph Purschke, Hinrich Schütze, Barbara Plank

Natural language processing (NLP) has largely focused on modelling standardized languages. More recently, attention has increasingly shifted to local, non-standardized languages and dialects. However, the relevant speake…

Machine Translation

MICE: A Crosslinguistic Emotion Corpus in Malay, Indonesian, Chinese and English

2021-06-09 · Ng Bee Chin, Yosephine Susanto, Erik Cambria

MICE is a corpus of emotion words in four languages which is currently working progress. There are two sections to this study, Part I: Emotion word corpus and Part II: Emotion word survey. In Part 1, the method of how th…

Survey

When Can Digital Personas Reliably Approximate Human Survey Findings?

2026-05-11 · Mumin Jia, Yilin Chen, Divya Sharma, Jairo Diaz-Rodriguez arxiv

Digital personas powered by Large Language Models (LLMs) are increasingly proposed as substitutes for human survey respondents, yet it remains unclear when they can reliably approximate human survey findings. We answer t…

Virtual Personas for Language Models via an Anthology of Backstories

2024-07-09 · Suhong Moon, Marwa Abdulhai, Minwoo Kang, Joseph Suh 외

Large language models (LLMs) are trained from vast repositories of text authored by millions of distinct authors, reflecting an enormous diversity of human traits. While these models bear the potential to be used as appr…

Diversity