paper-with-me

Papers

Morality, Machines and the Interpretation Problem: A Value-based, Wittgensteinian Approach to Building Moral Agents

2021-03-03 · Cosmin Badea, Gregory Artus

We present what we call the Interpretation Problem, whereby any rule in symbolic form is open to infinite interpretation in ways that we might disapprove of and argue that any attempt to build morality into machines is subject to it. We show how the Interpretation Problem in Artificial Intelligence is an illustration of Wittgenstein's general claim that no rule can contain the criteria for its own application, and that the risks created by this problem escalate in proportion to the degree to which to machine is causally connected to the world, in what we call the Law of Interpretative Exposure. Using game theory, we attempt to define the structure of normative spaces and argue that any rule-following within a normative space is guided by values that are external to that space and which cannot themselves be represented as rules. In light of this, we categorise the types of mistakes an artificial moral agent could make into Mistakes of Intention and Instrumental Mistakes, and we propose ways of building morality into machines by getting them to interpret the rules we give in accordance with these external values, through explicit moral reasoning, the Show, not Tell paradigm, the adjustment of causal power and structure of the agent, and relational values, with the ultimate aim that the machine develop a virtuous character and that the impact of the Interpretation Problem is minimised.

📄 PDF Abstract BibTeX arXiv:2103.02728

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Can Machines Learn Morality? The Delphi Experiment

2021-10-14 · Liwei Jiang, Jena D. Hwang, Chandra Bhagavatula, Ronan Le Bras 외

As AI systems become increasingly powerful and pervasive, there are growing concerns about machines' morality or a lack thereof. Yet, teaching morality to machines is a formidable task, as morality remains among the most…

DescriptiveEthics

The Wittgensteinian Representation Hypothesis: Is Language the Attractor of Multimodal Convergence?

2026-05-10 · Zhaoyang Zhang, Run Shao, Dongyue Wu, Jiajie Teng 외 arxiv

Understanding why independently trained neural networks from different modalities converge toward shared representations, and where this convergence leads, remains an open question in representation learning. All existin…

Representation LearningPoint Clouds

Knowledge Graphs meet Moral Values

2020-12-01 · Joint Conference on Lexical and Computational Semantics 2020 · Ioana Hulpu{\textcommabelow{s}}, Jonathan Kobbe, Heiner Stuckenschmidt, Graeme Hirst

Operationalizing morality is crucial for understanding multiple aspects of society that have moral values at their core {--} such as riots, mobilizing movements, public debates, etc. Moral Foundations Theory (MFT) has be…

Knowledge Graphs

Enhancing the Measurement of Social Effects by Capturing Morality

2019-06-01 · WS 2019 6 · Rezvaneh Rezapour, Saumil H. Shah, Jana Diesner

We investigate the relationship between basic principles of human morality and the expression of opinions in user-generated text data. We assume that people{'}s backgrounds, culture, and values are associated with their …

Cultural Vocal Bursts Intensity Prediction

MoralDial: A Framework to Train and Evaluate Moral Dialogue Systems via Moral Discussions

2022-12-21 · Hao Sun, Zhexin Zhang, Fei Mi, Yasheng Wang 외

Morality in dialogue systems has raised great attention in research recently. A moral dialogue system aligned with users' values could enhance conversation engagement and user connections. In this paper, we propose a fra…