paper-with-me

홈 › Papers

Can Large Language Models Anticipate Behavioral Responses to Social Policies? A Case of Pension Enrollment Prediction among China's Flexible Workers

2026-09-04 · Yumiao Li, Peixin Liu, Donglin Di, Chen Li, Runhuan Feng arxiv

Assessing the impacts of social policy changes is a widely acknowledged challenge for policymakers. Econometric methods can be unreliable when extrapolating to hypothetical scenarios, while field pilot programs are highly costly. In this paper, we propose using large language models (LLMs) as policy-assessment tools adapted from general-purpose models. We present FlexPension-LLM, the first domain-specialized large language model for a hierarchical pension-enrollment prediction task among flexible workers in China, and introduce DKI-RDistill, which injects policy-grounded cues into the prompt, including Probit-derived marginal effects and hukou-province pension rules. The method then uses LoRA/SFT to distill rationale-augmented supervision into an open-weight MoE student, with teacher errors corrected by regenerating those cases under ground-truth labels. On a CHFS 2019 blind split, FlexPension-LLM achieves 0.9316 Composite F1, surpassing its Claude Sonnet 4.5 teacher and 15 of 17 baselines, and is statistically indistinguishable from Claude Opus 4.6. Across four external surveys, it averages 0.7549 Composite F1 and shows the narrowest performance range among the strongest systems. Component analysis shows that gains come mainly from policy-grounded cue injection and error-filtered supervision, while rationales provide decision traces that can be checked against policy rules.

📄 PDF Abstract BibTeX arXiv:2609.05189

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

PHORECAST: Enabling AI Understanding of Public Health Outreach Across Populations

2025-10-02 · Rifaa Qadri, Anh Nhat Nhu, Swati Ramnath, Laura Yu Zheng 외 arxiv

Understanding how diverse individuals and communities respond to persuasive messaging holds significant potential for advancing personalized and socially aware machine learning. While Large Vision and Language Models (VL…

When simulations look right but causal effects go wrong: Large language models as behavioral simulators

2026-04-02 · Zonghan Li, Feng Ji arxiv

Behavioral simulation is increasingly used to anticipate responses to interventions. Large language models (LLMs) enable researchers to specify population characteristics and intervention context in natural language, but…

Infected Smallville: How Disease Threat Shapes Sociality in LLM Agents

2025-06-10 · Soyeon Choi, Kangwook Lee, Oliver Sng, Joshua M. Ackerman

How does the threat of infectious disease influence sociality among generative agents? We used generative agent-based modeling (GABM), powered by large language models, to experimentally test hypotheses about the behavio…

SCRAG: Social Computing-Based Retrieval Augmented Generation for Community Response Forecasting in Social Media Environments

2025-04-18 · Dachun Sun, You Lyu, Jinning Li, Yizhuo Chen 외

This paper introduces SCRAG, a prediction framework inspired by social computing, designed to forecast community responses to real or hypothetical social media posts. SCRAG can be used by public relations specialists (e.…

ArticlesPublic RelationsRAGRetrieval-augmented Generation

Improving Behavioral Alignment in LLM Social Simulations via Context Formation and Navigation

2026-01-04 · Letian Kong, Qianran, Jin, Renyu Zhang arxiv

Large language models (LLMs) are increasingly used to simulate human behavior in experimental settings, but they systematically diverge from human decisions in complex decision-making environments, where participants mus…