paper-with-me

Papers

Measuring an artificial intelligence agent's trust in humans using machine incentives

2022-12-27 · Tim Johnson, Nick Obradovich

Scientists and philosophers have debated whether humans can trust advanced artificial intelligence (AI) agents to respect humanity's best interests. Yet what about the reverse? Will advanced AI agents trust humans? Gauging an AI agent's trust in humans is challenging because--absent costs for dishonesty--such agents might respond falsely about their trust in humans. Here we present a method for incentivizing machine decisions without altering an AI agent's underlying algorithms or goal orientation. In two separate experiments, we then employ this method in hundreds of trust games between an AI agent (a Large Language Model (LLM) from OpenAI) and a human experimenter (author TJ). In our first experiment, we find that the AI agent decides to trust humans at higher rates when facing actual incentives than when making hypothetical decisions. Our second experiment replicates and extends these findings by automating game play and by homogenizing question wording. We again observe higher rates of trust when the AI agent faces real incentives. Across both experiments, the AI agent's trust decisions appear unrelated to the magnitude of stakes. Furthermore, to address the possibility that the AI agent's trust decisions reflect a preference for uncertainty, the experiments include two conditions that present the AI agent with a non-social decision task that provides the opportunity to choose a certain or uncertain option; in those conditions, the AI agent consistently chooses the certain option. Our experiments suggest that one of the most advanced AI language models to date alters its social behavior in response to incentives and displays behavior consistent with trust toward a human interlocutor when incentivized.

📄 PDF Abstract BibTeX arXiv:2212.13371

Code (0)

등록된 구현이 없습니다.

Tasks

AI AgentLanguage ModellingLarge Language Model

Similar Papers 제목 키워드 기반

Interactive embodied evolution for socially adept Artificial General Creatures

2024-07-31 · Kevin Godin-Dubois, Olivier Weissl, Karine Miras, Anna V. Kononova

We introduce here the concept of Artificial General Creatures (AGC) which encompasses "robotic or virtual agents with a wide enough range of capabilities to ensure their continued survival". With this in mind, we propose…

Ethics

VizTrust: A Visual Analytics Tool for Capturing User Trust Dynamics in Human-AI Communication

2025-03-10 · Xin Wang, Stephanie Tulk Jesso, Sadamori Kojaku, David M Neyens 외

Trust plays a fundamental role in shaping the willingness of users to engage and collaborate with artificial intelligence (AI) systems. Yet, measuring user trust remains challenging due to its complex and dynamic nature.…

An Approach to Joint Hybrid Decision Making between Humans and Artificial Intelligence

2025-11-29 · Jonas D. Rockbach, Sven Fuchs, Maren Bennewitz arxiv

Due to the progress in artificial intelligence, it is important to understand how capable artificial agents should be used when interacting with humans, since high level authority and responsibility often remain with the…

Decision Making

Cooperation in Human and Machine Agents: Promise Theory Considerations

2026-04-12 · M. Burgess arxiv

Agent based systems are more common than we may think. A Promise Theory perspective on cooperation, in systems of human-machine agents, offers a unified perspective on organization and functional design with semi-automat…

Dynamic Trust Calibration Using Contextual Bandits

2025-09-27 · Bruno M. Henrique, Eugene Santos arxiv

Trust calibration between humans and Artificial Intelligence (AI) is crucial for optimal decision-making in collaborative settings. Excessive trust can lead users to accept AI-generated outputs without question, overlook…