paper-with-me

홈 › Papers

Polistemics: Evaluating LLMs as Information Mediators in Politics & Elections

2026-07-28 · Baran Peters arxiv

As LLMs increasingly mediate the political information citizens rely on, there is still no standardized way to assess whether they do so responsibly. We introduce Polistemics, a theory-grounded benchmark for evaluating LLMs as mediators of political information in elections. Prior work has treated this task as reproduction rather than mediation, leaving its epistemic dimensions and interaction with imperfect information unaddressed. We ground the evaluation in Epistemic Modesty, a normative standard derived from citizens' epistemic agency, and test it across controlled settings that vary informational properties such as clarity, noise, and consistency. Applying the benchmark to three state-of-the-art LLMs on the 2025 German and Dutch elections, we find that high aggregate scores mask systematic failures. Models mediate reliably under clear evidence but break down under absent, vague, or contradictory information, while flattening the intensity of political language. These failures are likely driven by party priors, influenced by party labels and output language. Reliable mediation appears achievable, but no model delivers it consistently.

📄 PDF Abstract BibTeX arXiv:2607.25953

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators

2025-07-08 · Sungjib Lim, Woojung Song, Eun-Ju Lee, Yohan Jo arxiv

As psychometric surveys are increasingly used to assess the traits of large language models (LLMs), the need for scalable survey item generation suited for LLMs has also grown. A critical challenge here is ensuring the c…

SoCRATES: Towards Reliable Automated Evaluation of Proactive LLM Mediation across Domains and Socio-cognitive Variations

2026-06-04 · Taewon Yun, Hyeonseong Park, Jeonghwan Choi, Hayoon Park 외 arxiv

Evaluating LLM mediators remains challenging, as mediation unfolds as a real-time trajectory shaped by disputants' shifting emotions, intentions, and context. Existing testbeds rely on a few expert-authored domains, vary…

PoliticsBench: Benchmarking Political Values in Large Language Models with Multi-Turn Roleplay

2026-03-25 · Rohan Khetan, Ashna Khetan arxiv

While Large Language Models (LLMs) are increasingly used as primary sources of information, their potential for political bias may impact their objectivity. Existing benchmarks of LLM social bias primarily evaluate demog…

Large Language Models in Politics and Democracy: A Comprehensive Survey

2024-12-01 · Goshi Aoki

The advancement of generative AI, particularly large language models (LLMs), has a significant impact on politics and democracy, offering potential across various domains, including policymaking, political communication,…

Decision Making

Sensitivity Analysis for Causal Mediation through Text: an Application to Political Polarization

2021-11-01 · EMNLP (CINLP) 2021 11 · Graham Tierney, Alexander Volfovsky

We introduce a procedure to examine a text-as-mediator problem from a novel randomized experiment that studied the effect of conversations on political polarization. In this randomized experiment, Americans from the Demo…

Dimensionality ReductionSensitivity