paper-with-me

Papers

In Pursuit of Predictive Models of Human Preferences Toward AI Teammates

2025-01-31 · Ho Chit Siu, Jaime D. Peña, Yutai Zhou, Ross E. Allen

We seek measurable properties of AI agents that make them better or worse teammates from the subjective perspective of human collaborators. Our experiments use the cooperative card game Hanabi -- a common benchmark for AI-teaming research. We first evaluate AI agents on a set of objective metrics based on task performance, information theory, and game theory, which are measurable without human interaction. Next, we evaluate subjective human preferences toward AI teammates in a large-scale (N=241) human-AI teaming experiment. Finally, we correlate the AI-only objective metrics with the human subjective preferences. Our results refute common assumptions from prior literature on reinforcement learning, revealing new correlations between AI behaviors and human preferences. We find that the final game score a human-AI team achieves is less predictive of human preferences than esoteric measures of AI action diversity, strategic dominance, and ability to team with other AI. In the future, these correlations may help shape reward functions for training human-collaborative AI.

📄 PDF Abstract BibTeX arXiv:2503.15516

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Assisting Unknown Teammates in Unknown Tasks: Ad Hoc Teamwork under Partial Observability

2022-01-10 · João G. Ribeiro, Cassandro Martinho, Alberto Sardinha, Francisco S. Melo

In this paper, we present a novel Bayesian online prediction algorithm for the problem setting of ad hoc teamwork under partial observability (ATPO), which enables on-the-fly collaboration with unknown teammates performi…

HOLA-Drone: Hypergraphic Open-ended Learning for Zero-Shot Multi-Drone Cooperative Pursuit

2024-09-13 · Yang Li, Dengyu Zhang, Junfan Chen, Ying Wen 외

Zero-shot coordination (ZSC) is a significant challenge in multi-agent collaboration, aiming to develop agents that can coordinate with unseen partners they have not encountered before. Recent cutting-edge ZSC methods ha…

PADiff: Predictive and Adaptive Diffusion Policies for Ad Hoc Teamwork

2025-11-10 · Hohei Chan, Xinzhi Zhang, Antao Xiang, Weinan Zhang 외 arxiv

Ad hoc teamwork (AHT) requires agents to collaborate with previously unseen teammates, which is crucial for many real-world applications. The core challenge of AHT is to develop an ego agent that can predict and adapt to…

Emergent Behaviors in Multi-Agent Target Acquisition

2022-12-15 · Piyush K. Sharma, Erin Zaroukian, Derrik E. Asher, Bryson Howell

Only limited studies and superficial evaluations are available on agents' behaviors and roles within a Multi-Agent System (MAS). We simulate a MAS using Reinforcement Learning (RL) in a pursuit-evasion (a.k.a predator-pr…

Reinforcement Learning (RL)

AT-Drone: Benchmarking Adaptive Teaming in Multi-Drone Pursuit

2025-02-13 · Yang Li, Junfan Chen, Feng Xue, Jiabin Qiu 외

Adaptive teaming-the capability of agents to effectively collaborate with unfamiliar teammates without prior coordination-is widely explored in virtual video games but overlooked in real-world multi-robot contexts. Yet, …

BenchmarkingEdge-computing