paper-with-me

홈 › Papers

Every Act Has Its Price: Compressed Moral Composition in Frontier LLMs

2026-05-29 · Weijia Zhang, Ruiqi Chen, Yunze Xiao, Weihao Xuan arxiv

Existing LLM moral benchmarks usually ask which isolated moral act, value, or foundation a model prefers. This is useful but incomplete. Realistic judgments often require a model to combine several moral signals within the same option. We introduce Moral Trolley Arena, a two-stage blind ELO benchmark for measuring how LLMs compose moral evidence. The single-scene arena first calibrates individual moral acts from a 229-scenario corpus across five Moral Foundations Theory foundations; the composite arena then combines calibrated acts into two-act moral items over a controlled intensity grid and measures the resulting composite preferences. Across ten frontier models, composite judgments are largely predicted by component act strength, but the relation is consistently compressed rather than simply additive. Models also show non-additive intensity anchoring, bounded foundation-specific residuals after component control, and highly convergent composite preference surfaces across providers. These results suggest that moral audits should measure composition rules for moral evidence, not only rankings over isolated acts.

📄 PDF Abstract BibTeX arXiv:2606.11232

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Competing at Every Price Point with Agentic Evolution over a Menu of LLMs

2026-08-17 · Andrew Borthwick arxiv

Consider a firm that surveys its competition for a particular agentic task and seeks to offer superior accuracy at every competitor price point. A firm that Pareto-dominated its competitors would leave no rational custom…

Code Generation

HarvestBench: Measuring Whether LLM Agents Will Pay to Avoid Killing Animals

2026-09-03 · Jasmine Brazilek, Miles Tidmarsh, Matthias Endres, Anshuman Singh 외 hf

Benchmarks for the side effects an agent causes on the way to a goal already exist, but HarvestBench is the first to put a price on avoiding the side effect and to name that side effect as a living creature. It is a farm…

Reinforcement Learning

The yes-no bias of large language models reflects answer order and wording, not shifts in moral judgment

2026-07-06 · Haonan Huang arxiv

Large language models (LLMs) increasingly issue judgments read as binary verdicts, and a growing literature reports such judgments shifting under logically irrelevant changes of wording - among them an amplified yes-no b…

Triangle Fees

2023-06-29 · Rithvik Rao, Nihar Shah

Triangle fees are a novel fee structure for AMMs, in which marginal fees are decreasing in a trade's size. That decline is proportional to the movement in the AMM's implied price, i.e. for every basis point the trade mov…

Normative Robustness as a Frontier for Non-Verifiable Reasoning in LLMs

2026-06-10 · Elizaveta Tennant, Benjamin Henke, Anita Keshmirian, Murray Shanahan 외 arxiv

As LLMs increasingly serve in advisory and deliberative roles, users rely on them for non-verifiable reasoning in domains lacking objective ground truths. However, traditional evaluations of LLM reasoning focus almost ex…