paper-with-me

홈 › Papers

Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models

2024-02-26 · Paul Röttger, Valentin Hofmann, Valentina Pyatkin, Musashi Hinck, Hannah Rose Kirk, Hinrich Schütze, Dirk Hovy

Much recent work seeks to evaluate values and opinions in large language models (LLMs) using multiple-choice surveys and questionnaires. Most of this work is motivated by concerns around real-world LLM applications. For example, politically-biased LLMs may subtly influence society when they are used by millions of people. Such real-world concerns, however, stand in stark contrast to the artificiality of current evaluations: real users do not typically ask LLMs survey questions. Motivated by this discrepancy, we challenge the prevailing constrained evaluation paradigm for values and opinions in LLMs and explore more realistic unconstrained evaluations. As a case study, we focus on the popular Political Compass Test (PCT). In a systematic review, we find that most prior work using the PCT forces models to comply with the PCT's multiple-choice format. We show that models give substantively different answers when not forced; that answers change depending on how models are forced; and that answers lack paraphrase robustness. Then, we demonstrate that models give different answers yet again in a more realistic open-ended answer setting. We distill these findings into recommendations and open challenges in evaluating values and opinions in LLMs.

📄 PDF Abstract BibTeX arXiv:2402.16786

Code (1)

paul-rottger/llm-values-pct 공식 구현

Tasks

Multiple-choice

Methods 이 논문이 사용한 방법론

Focus 설명 없음
PCT 설명 없음

Similar Papers 제목 키워드 기반

Multilingual Political Views of Large Language Models: Identification and Steering

2025-07-30 · Daniil Gurgurov, Katharina Trinley, Ivan Vykopal, Josef van Genabith 외 arxiv

Large language models (LLMs) are increasingly used in everyday tools and applications, raising concerns about their potential influence on political views. While prior research has shown that LLMs often exhibit measurabl…

Modeling and analysis of a flexible spinning Euler-Bernoulli beam with centrifugal stiffening and softening: A Linear Fractional Representation approach with application to spinning spacecraft

2024-01-31 · Ricardo Rodrigues, Daniel Alazard, Francesco Sanfedino, Tommaso Mauriello 외

The derivation of a linear fractional representation (LFR) model for a flexible, spinning and uniform Euler-Bernoulli beam is accomplished using the {Lagrange} technique, fully capturing the centrifugal force generated b…

Cantilever Beam

Steering Towards Fairness: Mitigating Political Bias in LLMs

2025-08-12 · Afrozah Nadeem, Mark Dras, Usman Naseem arxiv

Recent advancements in large language models (LLMs) have enabled their widespread use across diverse real-world applications. However, concerns remain about their tendency to encode and reproduce ideological biases along…

Only a Little to the Left: A Theory-grounded Measure of Political Bias in Large Language Models

2025-03-20 · Mats Faulborn, Indira Sen, Max Pellert, Andreas Spitz 외

Prompt-based language models like GPT4 and LLaMa have been used for a wide variety of use cases such as simulating agents, searching for information, or for content analysis. For all of these applications and others, pol…

Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias

2026-01-08 · Adib Sakhawat, Tahsin Islam, Takia Farhin, Syed Rifat Raiyan 외 arxiv

As large language models (LLMs) are increasingly deployed, understanding how they express political positioning is important for evaluating alignment and downstream effects. We audit 26 contemporary LLMs using three poli…