paper-with-me

홈 › Papers

Exploring Persona-dependent LLM Alignment for the Moral Machine Experiment

2025-04-15 · Jiseon Kim, Jea Kwon, Luiz Felipe Vecchietti, Alice Oh, Meeyoung Cha

Deploying large language models (LLMs) with agency in real-world applications raises critical questions about how these models will behave. In particular, how will their decisions align with humans when faced with moral dilemmas? This study examines the alignment between LLM-driven decisions and human judgment in various contexts of the moral machine experiment, including personas reflecting different sociodemographics. We find that the moral decisions of LLMs vary substantially by persona, showing greater shifts in moral decisions for critical tasks than humans. Our data also indicate an interesting partisan sorting phenomenon, where political persona predominates the direction and degree of LLM decisions. We discuss the ethical implications and risks associated with deploying these models in applications that involve moral decisions.

📄 PDF Abstract BibTeX arXiv:2504.10886

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions

2026-04-23 · Jiseon Kim, Jea Kwon, Luiz Felipe Vecchietti, Wenchao Dong 외 arxiv

Human moral judgment is context-dependent and modulated by interpersonal relationships. As large language models (LLMs) increasingly function as decision-support systems, determining whether they encode these social nuan…

The Moral Turing Test: Evaluating Human-LLM Alignment in Moral Decision-Making

2024-10-09 · Basile Garcia, Crystal Qian, Stefano Palminteri

As large language models (LLMs) become increasingly integrated into society, their alignment with human morals is crucial. To better understand this alignment, we created a large corpus of human- and LLM-generated respon…

Decision MakingMoral Scenarios

Culturally Grounded Personas in Large Language Models: Characterization and Alignment with Socio-Psychological Value Frameworks

2026-01-29 · Candida M. Greco, Lucio La Cava, Andrea Tagarelli arxiv

Despite the growing utility of Large Language Models (LLMs) for simulating human behavior, the extent to which these synthetic personas accurately reflect world and moral value systems across different cultural condition…

Towards "Differential AI Psychology" and in-context Value-driven Statement Alignment with Moral Foundations Theory

2024-08-21 · Simon Münker

Contemporary research in social sciences is increasingly utilizing state-of-the-art statistical language models to annotate or generate content. While these models perform benchmark-leading on common language tasks and s…

Language ModellingSurvey

MORALISE: A Structured Benchmark for Moral Alignment in Visual Language Models

2025-05-20 · Xiao Lin, Zhining Liu, Ze Yang, Gaotang Li 외

Warning: This paper contains examples of harmful language and images. Reader discretion is advised. Recently, vision-language models have demonstrated increasing influence in morally sensitive domains such as autonomous …

Autonomous DrivingMultimodal Reasoning