paper-with-me

Papers

Kantian Deontology Meets AI Alignment: Towards Morally Grounded Fairness Metrics

2023-11-09 · Carlos Mougan, Joshua Brand

Deontological ethics, specifically understood through Immanuel Kant, provides a moral framework that emphasizes the importance of duties and principles, rather than the consequences of action. Understanding that despite the prominence of deontology, it is currently an overlooked approach in fairness metrics, this paper explores the compatibility of a Kantian deontological framework in fairness metrics, part of the AI alignment field. We revisit Kant's critique of utilitarianism, which is the primary approach in AI fairness metrics and argue that fairness principles should align with the Kantian deontological framework. By integrating Kantian ethics into AI alignment, we not only bring in a widely-accepted prominent moral theory but also strive for a more morally grounded AI landscape that better balances outcomes and procedures in pursuit of fairness and justice.

📄 PDF Abstract BibTeX arXiv:2311.05227

Code (0)

등록된 구현이 없습니다.

Tasks

Action UnderstandingEthicsFairness

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Game-theoretic Models of Moral and Other-Regarding Agents

2020-12-17 · Gabriel Istrate

We investigate Kantian equilibria in finite normal form games, a class of non-Nashian, morally motivated courses of action that was recently proposed in the economics literature. We highlight a number of problems with su…

Form

Game-Theoretic Models of Moral and Other-Regarding Agents (extended abstract)

2021-06-22 · Gabriel Istrate

We investigate Kantian equilibria in finite normal form games, a class of non-Nashian, morally motivated courses of action that was recently proposed in the economics literature. We highlight a number of problems with su…

Form

Doing the right thing (or not) in a lemons-like situation: on the role of social preferences and Kantian moral concerns

2024-05-21 · Ingela Alger, José Ignacio Rivero-Wildemauwe

We conduct a laboratory experiment using framing to assess the willingness to ``sell a lemon'', i.e., to undertake an action that benefits self but hurts the other (the ``buyer''). We seek to disentangle the role of othe…

Bounded Morality: Defining the Space of Moral Computation

2026-04-01 · Max Kanwal, Caryn Tran, Patrick Mineault arxiv

Moral cognition has traditionally been modeled as adherence to fixed ethical theories--deontology, consequentialism, virtue ethics--implemented as static rules or value functions. We propose Bounded Morality, a formal fr…

Hybrid Approaches for Moral Value Alignment in AI Agents: a Manifesto

2023-12-04 · Elizaveta Tennant, Stephen Hailes, Mirco Musolesi

Increasing interest in ensuring the safety of next-generation Artificial Intelligence (AI) systems calls for novel approaches to embedding morality into autonomous agents. This goal differs qualitatively from traditional…

Ethics