paper-with-me

Papers

Normative Evaluation of Large Language Models with Everyday Moral Dilemmas

2025-01-30 · Pratik S. Sachdeva, Tom van Nuenen

The rapid adoption of large language models (LLMs) has spurred extensive research into their encoded moral norms and decision-making processes. Much of this research relies on prompting LLMs with survey-style questions to assess how well models are aligned with certain demographic groups, moral beliefs, or political ideologies. While informative, the adherence of these approaches to relatively superficial constructs tends to oversimplify the complexity and nuance underlying everyday moral dilemmas. We argue that auditing LLMs along more detailed axes of human interaction is of paramount importance to better assess the degree to which they may impact human beliefs and actions. To this end, we evaluate LLMs on complex, everyday moral dilemmas sourced from the "Am I the Asshole" (AITA) community on Reddit, where users seek moral judgments on everyday conflicts from other community members. We prompted seven LLMs to assign blame and provide explanations for over 10,000 AITA moral dilemmas. We then compared the LLMs' judgments and explanations to those of Redditors and to each other, aiming to uncover patterns in their moral reasoning. Our results demonstrate that large language models exhibit distinct patterns of moral judgment, varying substantially from human evaluations on the AITA subreddit. LLMs demonstrate moderate to high self-consistency but low inter-model agreement. Further analysis of model explanations reveals distinct patterns in how models invoke various moral principles. These findings highlight the complexity of implementing consistent moral reasoning in artificial systems and the need for careful evaluation of how different models approach ethical judgment. As LLMs continue to be used in roles requiring ethical decision-making such as therapists and companions, careful evaluation is crucial to mitigate potential biases and limitations.

📄 PDF Abstract BibTeX arXiv:2501.18081

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

EMNLP: Educator-role Moral and Normative Large Language Models Profiling

2025-08-21 · Yilin Jiang, Mingzi Zhang, Sheng Jin, Zengyi Yu 외 arxiv

Simulating Professions (SP) enables Large Language Models (LLMs) to emulate professional roles. However, comprehensive psychological and ethical evaluation in these contexts remains lacking. This paper introduces EMNLP, …

M$^3$oralBench: A MultiModal Moral Benchmark for LVLMs

2024-12-30 · Bei Yan, Jie Zhang, ZhiYuan Chen, Shiguang Shan 외

Recently, large foundation models, including large language models (LLMs) and large vision-language models (LVLMs), have become essential tools in critical fields such as law, finance, and healthcare. As these models inc…

Moral Scenarios

Unpacking the Ethical Value Alignment in Big Models

2023-10-26 · Xiaoyuan Yi, Jing Yao, Xiting Wang, Xing Xie

Big models have greatly advanced AI's ability to understand, generate, and manipulate information and content, enabling numerous applications. However, as these models become increasingly integrated into everyday life, t…

Ethics

Differences in the Moral Foundations of Large Language Models

2025-11-14 · Peter Kirgis arxiv

Large language models are increasingly being used in critical domains of politics, business, and education, but the nature of their normative ethical judgment remains opaque. Alignment research has, to date, not sufficie…

Towards Theory-based Moral AI: Moral AI with Aggregating Models Based on Normative Ethical Theory

2023-06-20 · Masashi Takeshita, Rzepka Rafal, Kenji Araki

Moral AI has been studied in the fields of philosophy and artificial intelligence. Although most existing studies are only theoretical, recent developments in AI have made it increasingly necessary to implement AI with m…

EthicsPhilosophy