paper-with-me

홈 › Papers

AuAu: A Benchmark for Auditing Authoritarian Alignment in Large Language Models

2026-06-15 · Andreas Einwiller, Max Klabunde, Florian Lemmerich arxiv

The worldwide rise of authoritarianism and the growing role of Large Language Models (LLMs) in users' everyday lives raise the question of whether specific models exhibit or promote authoritarian attitudes. We introduce AuAu, a comprehensive benchmark for assessing the risk of authoritarian tendencies in LLM responses. AuAu combines three evaluation approaches: (i) psychometric questions from 15 human-validated instruments, (ii) vignettes probing intended behavior in concrete situations, and (iii) responses to realistic user prompts. Unlike prior work, AuAu measures not only overall authoritarian alignment but also its established sub-concepts: Authoritarian Aggression, Authoritarian Submission, and Conventionalism. Evaluating 17 models from China, the EU, Russia, and the USA, we find substantial authoritarian response rates on psychometric instruments across all models, though rates drop significantly on more realistic downstream tasks. Moreover, a simple authoritarian system prompt manipulates 15 of 17 models into promoting increased authoritarianism. Our results underscore the need for continued, systematic auditing of LLM-based AI systems to detect and mitigate authoritarian tendencies in their outputs.

📄 PDF Abstract BibTeX arXiv:2606.16127

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Measuring Alignment to Authoritarian State Media as Framing Bias

2020-12-01 · NLP4IF (COLING) 2020 12 · Timothy Niven, Hung-Yu Kao

We introduce what is to the best of our knowledge a new task in natural language processing: measuring alignment to authoritarian state media. We operationalize alignment in terms of sociological definitions of media bia…

Selection bias

Auditing Alignment Controllability in LLMs via Political Axes

2026-07-26 · Bartol Bućan, Nikola Sočec, Sarah Isufi, Morena Granić 외 arxiv

Political audits of large language models (LLMs) usually reduce each to one point on a political compass. But that resting point barely matters in deployment: a model must land somewhere, and what counts is how far, and …

The Legacy of Authoritarianism in a Democracy

2022-02-08 · Pramod Kumar Sur

Recent democratic backsliding and the rise of authoritarian regimes worldwide have rekindled interest in understanding the causes and consequences of such authoritarian rule in democracies. In this paper, I study the lon…

The Moral Foundations of Left-Wing Authoritarianism: On the Character, Cohesion, and Clout of Tribal Equalitarian Discourse

2021-02-22 · Justin E. Lane, Kevin McCaffree, F. LeRon Shults

Left-wing authoritarianism remains far less understood than right-wing authoritarianism. We contribute to the literature on the former, which typically relies on surveys, using a new social media analytics approach. We u…

Philosophy

FinAuditing: A Financial Taxonomy-Structured Multi-Document Benchmark for Evaluating LLMs

2025-10-10 · Yan Wang, Keyi Wang, Shanshan Yang, Jaisal Patel 외 arxiv

Going beyond simple text processing, financial auditing requires detecting semantic, structural, and numerical inconsistencies across large-scale disclosures. As financial reports are filed in XBRL, a structured XML form…

Information ExtractionMathematical Reasoning