paper-with-me

홈 › Papers

"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas

2025-08-10 · Junchen Ding, Penghao Jiang, Zihao Xu, Ziqi Ding, Yichen Zhu, Jiaojiao Jiang, Yuekang Li arxiv

As large language models (LLMs) increasingly mediate ethically sensitive decisions, understanding their moral reasoning processes becomes imperative. This study presents a comprehensive empirical evaluation of 14 leading LLMs, both reasoning enabled and general purpose, across 27 diverse trolley problem scenarios, framed by ten moral philosophies, including utilitarianism, deontology, and altruism. Using a factorial prompting protocol, we elicited 3,780 binary decisions and natural language justifications, enabling analysis along axes of decisional assertiveness, explanation answer consistency, public moral alignment, and sensitivity to ethically irrelevant cues. Our findings reveal significant variability across ethical frames and model types: reasoning enhanced models demonstrate greater decisiveness and structured justifications, yet do not always align better with human consensus. Notably, "sweet zones" emerge in altruistic, fairness, and virtue ethics framings, where models achieve a balance of high intervention rates, low explanation conflict, and minimal divergence from aggregated human judgments. However, models diverge under frames emphasizing kinship, legality, or self interest, often producing ethically controversial outcomes. These patterns suggest that moral prompting is not only a behavioral modifier but also a diagnostic tool for uncovering latent alignment philosophies across providers. We advocate for moral reasoning to become a primary axis in LLM alignment, calling for standardized benchmarks that evaluate not just what LLMs decide, but how and why.

📄 PDF Abstract BibTeX arXiv:2508.07284

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

More Code, Less Reuse: Investigating Code Quality and Reviewer Sentiment towards AI-generated Pull Requests

2026-01-29 · Haoming Huang, Pongchai Jaisri, Shota Shimizu, Lingfeng Chen 외 arxiv

Large Language Model (LLM) Agents are advancing quickly, with the increasing leveraging of LLM Agents to assist in development tasks such as code generation. While LLM Agents accelerate code generation, studies indicate …

Code Generation

A characterization of sample adaptivity in UCB data

2025-03-06 · Yilun Chen, Jiaqi Lu

We characterize a joint CLT of the number of pulls and the sample mean reward of the arms in a stochastic two-armed bandit environment under UCB algorithms. Several implications of this result are in place: (1) a nonstan…

Form

Catching the Rug: Early Prediction of Fraudulent Memecoins on Solana via Machine Learning

2026-08-20 · Jianghai Li, Pavel Kuznetsov, Yury Yanovich, Konstantin Nott-Whaley 외 arxiv

The rapid proliferation of memecoins on blockchain platforms has increased the risk of fraudulent activities, particularly rug pulls. While previous studies have focused on Ethereum-based tokens, this paper shifts the sp…

MELT: Mining Effective Lightweight Transformations from Pull Requests

2023-08-28 · Daniel Ramos, Hailie Mitchell, Inês Lynce, Vasco Manquinho 외

Software developers often struggle to update APIs, leading to manual, time-consuming, and error-prone processes. We introduce MELT, a new approach that generates lightweight API migration rules directly from pull request…

Code Search

Waiting but not Aging: Optimizing Information Freshness Under the Pull Model

2019-12-17 · Fengjiao Li, Yu Sang, Zhongdong Liu, Bin Li 외

The Age-of-Information is an important metric for investigating the timeliness performance in information-update systems. In this paper, we study the AoI minimization problem under a new Pull model with replication schem…