Kantian Deontology Meets AI Alignment: Towards Morally Grounded Fairness Metrics
Deontological ethics, specifically understood through Immanuel Kant, provides a moral framework that emphasizes the importance of duties and principles, rather than the consequences of action. Understanding that despite the prominence of deontology, it is currently an overlooked approach in fairness metrics, this paper explores the compatibility of a Kantian deontological framework in fairness metrics, part of the AI alignment field. We revisit Kant's critique of utilitarianism, which is the primary approach in AI fairness metrics and argue that fairness principles should align with the Kantian deontological framework. By integrating Kantian ethics into AI alignment, we not only bring in a widely-accepted prominent moral theory but also strive for a more morally grounded AI landscape that better balances outcomes and procedures in pursuit of fairness and justice.
Code (0)
등록된 구현이 없습니다.
Tasks
Action UnderstandingEthicsFairnessMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Game-theoretic Models of Moral and Other-Regarding Agents
We investigate Kantian equilibria in finite normal form games, a class of non-Nashian, morally motivated courses of action that was recently proposed in the economics literature. We highlight a number of problems with su…
FormGame-Theoretic Models of Moral and Other-Regarding Agents (extended abstract)
We investigate Kantian equilibria in finite normal form games, a class of non-Nashian, morally motivated courses of action that was recently proposed in the economics literature. We highlight a number of problems with su…
FormDoing the right thing (or not) in a lemons-like situation: on the role of social preferences and Kantian moral concerns
We conduct a laboratory experiment using framing to assess the willingness to ``sell a lemon'', i.e., to undertake an action that benefits self but hurts the other (the ``buyer''). We seek to disentangle the role of othe…
Bounded Morality: Defining the Space of Moral Computation
Moral cognition has traditionally been modeled as adherence to fixed ethical theories--deontology, consequentialism, virtue ethics--implemented as static rules or value functions. We propose Bounded Morality, a formal fr…
Hybrid Approaches for Moral Value Alignment in AI Agents: a Manifesto
Increasing interest in ensuring the safety of next-generation Artificial Intelligence (AI) systems calls for novel approaches to embedding morality into autonomous agents. This goal differs qualitatively from traditional…
Ethics