paper-with-me

홈 › Papers

Fairness and Deception in Human Interactions with Artificial Agents

2023-12-06 · Theodor Cimpeanu, Alexander J. Stewart

Online information ecosystems are now central to our everyday social interactions. Of the many opportunities and challenges this presents, the capacity for artificial agents to shape individual and collective human decision-making in such environments is of particular importance. In order to assess and manage the impact of artificial agents on human well-being, we must consider not only the technical capabilities of such agents, but the impact they have on human social dynamics at the individual and population level. We approach this problem by modelling the potential for artificial agents to "nudge" attitudes to fairness and cooperation in populations of human agents, who update their behavior according to a process of social learning. We show that the presence of artificial agents in a population playing the ultimatum game generates highly divergent, multi-stable outcomes in the learning dynamics of human agents' behaviour. These outcomes correspond to universal fairness (successful nudging), universal selfishness (failed nudging), and a strategy of fairness towards artificial agents and selfishness towards other human agents (unintended consequences of nudging). We then consider the consequences of human agents shifting their behavior when they are aware that they are interacting with an artificial agent. We show that under a wide range of circumstances artificial agents can achieve optimal outcomes in their interactions with human agents while avoiding deception. However we also find that, in the donation game, deception tends to make nudging easier to achieve.

📄 PDF Abstract BibTeX arXiv:2312.03645

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingFairness

Methods 이 논문이 사용한 방법론

AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…

Similar Papers 제목 키워드 기반

OpenDeception: Benchmarking and Investigating AI Deceptive Behaviors via Open-ended Interaction Simulation

2025-04-18 · Yichen Wu, Xudong Pan, Geng Hong, Min Yang

As the general capabilities of large language models (LLMs) improve and agent applications become more widespread, the underlying deception risks urgently require systematic evaluation and effective oversight. Unlike exi…

Benchmarking

Lies We Can See: Joint Verbal and Non-Verbal Deception by VLM Agents in Embodied Social Interactions

2026-08-31 · Jaewoo Ahn, Junseo Kim, Hyunseo Kim, Heeseung Yun 외 arxiv

Strategic deception by LLM and VLM agents has emerged as a central AI alignment and safety concern. Social-deduction games (where each player holds a hidden role and communicates with others to deduce identities) serve a…

Deception in Reinforced Autonomous Agents

2024-05-07 · Atharvan Dogra, Krishna Pillutla, Ameet Deshpande, Ananya B Sai 외

We explore the ability of large language model (LLM)-based agents to engage in subtle deception such as strategically phrasing and intentionally manipulating information to misguide and deceive other agents. This harmful…

Deception DetectionHallucinationLanguage ModelingLanguage Modelling+2

Deception Analysis with Artificial Intelligence: An Interdisciplinary Perspective

2024-06-09 · Stefan Sarkadi

Humans and machines interact more frequently than ever and our societies are becoming increasingly hybrid. A consequence of this hybridisation is the degradation of societal trust due to the prevalence of AI-enabled dece…

EthicsPhilosophy

Deceptive Games

2018-01-31 · Damien Anderson, Matthew Stephenson, Julian Togelius, Christian Salge 외

Deceptive games are games where the reward structure or other aspects of the game are designed to lead the agent away from a globally optimal policy. While many games are already deceptive to some extent, we designed a s…