paper-with-me

홈 › Papers

In Trust We Survive: Emergent Trust Learning

2026-03-18 · Qianpu Chen, Giulio Barbero, Mike Preuss, Derya Soydaner arxiv

We introduce Emergent Trust Learning (ETL), a lightweight, trust-based control algorithm that can be plugged into existing AI agents. It enables these to reach cooperation in competitive game environments under shared resources. Each agent maintains a compact internal trust state, which modulates memory, exploration, and action selection. ETL requires only individual rewards and local observations and incurs negligible computational and communication overhead. We evaluate ETL in three environments: In a grid-based resource world, trust-based agents reduce conflicts and prevent long-term resource depletion while achieving competitive individual returns. In a hierarchical Tower environment with strong social dilemmas and randomised floor assignments, ETL sustains high survival rates and recovers cooperation even after extended phases of enforced greed. In the Iterated Prisoner's Dilemma, the algorithm generalises to a strategic meta-game, maintaining cooperation with reciprocal opponents while avoiding long-term exploitation by defectors. Code will be released upon publication.

📄 PDF Abstract BibTeX arXiv:2603.17564

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Trust and Its Betrayal under Three Representational Strategies

2026-07-31 · Mihnea C. Moldoveanu, Joel A. C. Baum arxiv

Trust is a propositional attitude of a distinctive kind: to trust is to rely on another under conditions where reliance could be disappointed, and the disappointment of trust---betrayal---differs qualitatively from the d…

Trust-Based Social Learning for Communication (TSLEC) Protocol Evolution in Multi-Agent Reinforcement Learning

2025-11-24 · Abraham Itzhak Weinberg arxiv

Emergent communication in multi-agent systems typically occurs through independent learning, resulting in slow convergence and potentially suboptimal protocols. We introduce TSLEC (Trust-Based Social Learning with Emerge…

Multi-agent Reinforcement Learning

Exploratory Models of Human-AI Teams: Leveraging Human Digital Twins to Investigate Trust Development

2024-11-01 · Daniel Nguyen, Myke C. Cohen, Hsien-Te Kao, Grant Engberson 외

As human-agent teaming (HAT) research continues to grow, computational methods for modeling HAT behaviors and measuring HAT effectiveness also continue to develop. One rising method involves the use of human digital twin…

Experimental Design

Converging Measures and an Emergent Model: A Meta-Analysis of Human-Automation Trust Questionnaires

2023-03-24 · Yosef S. Razin, Karen M. Feigh

A significant challenge to measuring human-automation trust is the amount of construct proliferation, models, and questionnaires with highly variable validation. However, all agree that trust is a crucial element of tech…

Common Sense ReasoningSurvey

Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages

2026-07-06 · Lucas Pinto arxiv

Trusted monitoring is a central defense in AI control: a cheaper trusted model scores an untrusted model's actions for sabotage, and the most suspicious are audited or deferred. Such monitors are evaluated against one or…