paper-with-me

Papers

Grounding Value Alignment with Ethical Principles

2019-07-11 · Tae Wan Kim, Thomas Donaldson, John Hooker

An important step in the development of value alignment (VA) systems in AI is understanding how values can interrelate with facts. Designers of future VA systems will need to utilize a hybrid approach in which ethical reasoning and empirical observation interrelate successfully in machine behavior. In this article we identify two problems about this interrelation that have been overlooked by AI discussants and designers. The first problem is that many AI designers commit inadvertently a version of what has been called by moral philosophers the "naturalistic fallacy," that is, they attempt to derive an "ought" from an "is." We illustrate when and why this occurs. The second problem is that AI designers adopt training routines that fail fully to simulate human ethical reasoning in the integration of ethical principles and facts. Using concepts of quantified modal logic, we proceed to offer an approach that promises to simulate ethical reasoning in humans by connecting ethical principles on the one hand and propositions about states of affairs on the other.

📄 PDF Abstract BibTeX arXiv:1907.05447

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Taking Principles Seriously: A Hybrid Approach to Value Alignment

2020-12-21 · Tae Wan Kim, John Hooker, Thomas Donaldson

An important step in the development of value alignment (VA) systems in AI is understanding how VA can reflect valid ethical principles. We propose that designers of VA systems incorporate ethics by utilizing a hybrid ap…

Ethicsvalid

Wide Reflective Equilibrium in LLM Alignment: Bridging Moral Epistemology and AI Safety

2025-05-31 · Matthew Brophy

As large language models (LLMs) become more powerful and pervasive across society, ensuring these systems are beneficial, safe, and aligned with human values is crucial. Current alignment techniques, like Constitutional …

Ethical Reasoning over Moral Alignment: A Case and Framework for In-Context Ethical Policies in LLMs

2023-10-11 · Abhinav Rao, Aditi Khandelwal, Kumar Tanmay, Utkarsh Agarwal 외

In this position paper, we argue that instead of morally aligning LLMs to specific set of ethical principles, we should infuse generic ethical reasoning capabilities into them so that they can handle value pluralism at a…

EthicsPosition

Denevil: Towards Deciphering and Navigating the Ethical Values of Large Language Models via Instruction Learning

2023-10-17 · Shitong Duan, Xiaoyuan Yi, Peng Zhang, Tun Lu 외

Large Language Models (LLMs) have made unprecedented breakthroughs, yet their increasing integration into everyday life might raise societal risks due to generated unethical content. Despite extensive study on specific i…

EthicsPhilosophy

TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

2023-06-20 · Yue Huang, Qihui Zhang, Philip S. Y, Lichao Sun

Large Language Models (LLMs) such as ChatGPT, have gained significant attention due to their impressive natural language processing capabilities. It is crucial to prioritize human-centered principles when utilizing these…