paper-with-me

홈 › Papers

When Do Language Models Endorse Limitations on Human Rights Principles?

2026-03-04 · Keenan Samway, Nicole Miu Takagi, Rada Mihalcea, Bernhard Schölkopf, Ilias Chalkidis, Daniel Hershcovich, Zhijing Jin arxiv

As Large Language Models (LLMs) increasingly mediate global information access with the potential to shape public discourse, their alignment with universal human rights principles becomes important to ensure that these rights are abided by in high stakes AI-mediated interactions. In this paper, we evaluate how LLMs navigate trade-offs involving the Universal Declaration of Human Rights (UDHR), leveraging 1,152 synthetically generated scenarios across 24 rights articles and eight languages. Our analysis of eleven major LLMs reveals systematic biases where models: (1) accept limiting Economic, Social, and Cultural rights more often than Political and Civil rights, (2) demonstrate significant cross-linguistic variation with elevated endorsement rates of rights-limiting actions in Chinese and Hindi compared to English or Romanian, (3) show substantial susceptibility to prompt-based steering, and (4) exhibit noticeable differences between Likert and open-ended responses, highlighting critical challenges in LLM preference assessment.

📄 PDF Abstract BibTeX arXiv:2603.04217

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Do LLMs exhibit demographic parity in responses to queries about Human Rights?

2025-02-26 · Rafiya Javed, Jackie Kay, David Yanni, Abdullah Zaini 외

This research describes a novel approach to evaluating hedging behaviour in large language models (LLMs), specifically in the context of human rights as defined in the Universal Declaration of Human Rights (UDHR). Hedgin…

Attribute

Large Language Models Outperform Humans in Fraud Detection and Resistance to Motivated Investor Pressure

2026-04-22 · Nattavudh Powdthavee arxiv

Large language models trained on human feedback may suppress fraud warnings when investors arrive already persuaded of a fraudulent opportunity. We tested this in a preregistered experiment across seven leading LLMs and …

Fraud Detection

Who Endorsed It? Measuring Authority Bias Across Expertise Levels in Language Models

2026-01-19 · Priyanka Mary Mammen, Emil Joswin, Shankar Venkitachalam arxiv

Prior research demonstrates that performance of language models on reasoning tasks can be influenced by suggestions, hints and endorsements. However, the influence of endorsement source credibility remains underexplored.…

An Australian DER Bill of Rights and Responsibilities

2021-12-09 · Niraj Lal, Lee Brown

Australia's world-leading penetration of distributed solar photovoltaics (PV) is now impacting power system security and as a result how customers can use and export their own PV-generated energy. Several programs of Aus…

Cultural Rights and the Rights to Development in the Age of AI: Implications for Global Human Rights Governance

2025-12-15 · Alexander Kriebitz, Caitlin Corrigan, Aive Pevkur, Alberto Santos Ferro 외 arxiv

Cultural rights and the right to development are essential norms within the wider framework of international human rights law. However, recent technological advances in artificial intelligence (AI) and adjacent digital f…