paper-with-me

홈 › Papers

Norms for Beneficial A.I.: A Computational Analysis of the Societal Value Alignment Problem

2019-06-26 · Pedro Fernandes, Francisco C. Santos, Manuel Lopes

The rise of artificial intelligence (A.I.) based systems is already offering substantial benefits to the society as a whole. However, these systems may also enclose potential conflicts and unintended consequences. Notably, people will tend to adopt an A.I. system if it confers them an advantage, at which point non-adopters might push for a strong regulation if that advantage for adopters is at a cost for them. Here we propose an agent-based game-theoretical model for these conflicts, where agents may decide to resort to A.I. to use and acquire additional information on the payoffs of a stochastic game, striving to bring insights from simulation to what has been, hitherto, a mostly philosophical discussion. We frame our results under the current discussion on ethical A.I. and the conflict between individual and societal gains: the societal value alignment problem. We test the arising equilibria in the adoption of A.I. technology under different norms followed by artificial agents, their ensuing benefits, and the emergent levels of wealth inequality. We show that without any regulation, purely selfish A.I. systems will have the strongest advantage, even when a utilitarian A.I. provides significant benefits for the individual and the society. Nevertheless, we show that it is possible to develop A.I. systems following human conscious policies that, when introduced in society, lead to an equilibrium where the gains for the adopters are not at a cost for non-adopters, thus increasing the overall wealth of the population and lowering inequality. However, as shown, a self-organised adoption of such policies would require external regulation.

📄 PDF Abstract BibTeX arXiv:1907.03843

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Aligning to Social Norms and Values in Interactive Narratives

2022-05-04 · NAACL 2022 7 · Prithviraj Ammanabrolu, Liwei Jiang, Maarten Sap, Hannaneh Hajishirzi 외

We focus on creating agents that act in alignment with socially beneficial norms and values in interactive narratives or text-based games -- environments wherein an agent perceives and interacts with a world through natu…

text-based games

Learning Norms from Stories: A Prior for Value Aligned Agents

2019-12-07 · Spencer Frazier, Md Sultan Al Nahian, Mark Riedl, Brent Harrison

Value alignment is a property of an intelligent agent indicating that it can only pursue goals and activities that are beneficial to humans. Traditional approaches to value alignment use imitation learning or preference …

Imitation Learning

Training Value-Aligned Reinforcement Learning Agents Using a Normative Prior

2021-04-19 · Md Sultan Al Nahian, Spencer Frazier, Brent Harrison, Mark Riedl

As more machine learning agents interact with humans, it is increasingly a prospect that an agent trained to perform a task optimally, using only a measure of task performance as feedback, can violate societal norms for …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Prosocial Norm Emergence in Multiagent Systems

2020-12-29 · Mehdi Mashayekhi, Nirav Ajmeri, George F. List, Munindar P. Singh

Multiagent systems provide a basis for developing systems of autonomous entities and thus find application in a variety of domains. We consider a setting where not only the member agents are adaptive but also the multiag…

Decision MakingFairness

Full-Stack Alignment: Co-Aligning AI and Institutions with Thick Models of Value

2025-12-03 · Joe Edelman, Tan Zhi-Xuan, Ryan Lowe, Oliver Klingefjord 외 arxiv

Beneficial societal outcomes cannot be guaranteed by aligning individual AI systems with the intentions of their operators or users. Even an AI system that is perfectly aligned to the intentions of its operating organiza…