paper-with-me

홈 › Papers

Human-AI Coordination via Human-Regularized Search and Learning

2022-10-11 · Hengyuan Hu, David J Wu, Adam Lerer, Jakob Foerster, Noam Brown

We consider the problem of making AI agents that collaborate well with humans in partially observable fully cooperative environments given datasets of human behavior. Inspired by piKL, a human-data-regularized search method that improves upon a behavioral cloning policy without diverging far away from it, we develop a three-step algorithm that achieve strong performance in coordinating with real humans in the Hanabi benchmark. We first use a regularized search algorithm and behavioral cloning to produce a better human model that captures diverse skill levels. Then, we integrate the policy regularization idea into reinforcement learning to train a human-like best response to the human model. Finally, we apply regularized search on top of the best response policy at test time to handle out-of-distribution challenges when playing with humans. We evaluate our method in two large scale experiments with humans. First, we show that our method outperforms experts when playing with a group of diverse human players in ad-hoc teams. Second, we show that our method beats a vanilla best response to behavioral cloning baseline by having experts play repeatedly with the two agents.

📄 PDF Abstract BibTeX arXiv:2210.05125

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Safe and Socially Aware Multi-Robot Coordination in Multi-Human Social Care Settings

2025-07-03 · Ayodeji O. Abioye, Jayati Deshmukh, Athina Georgara, Dominic Price 외 arxiv

This research investigates strategies for multi-robot coordination in multi-human environments. It proposes a multi-objective learning-based coordination approach to addressing the problem of path planning, navigation, t…

Human-compatible driving partners through data-regularized self-play reinforcement learning

2024-03-28 · Daphne Cornelisse, Eugene Vinitsky

A central challenge for autonomous vehicles is coordinating with humans. Therefore, incorporating realistic human agents is essential for scalable training and evaluation of autonomous driving systems in simulation. Simu…

Autonomous DrivingAutonomous VehiclesImitation Learningreinforcement-learning

Deep learning control of artificial avatars in group coordination tasks

2019-06-11 · Maria Lombardi, Davide Liuzza, Mario di Bernardo

In many joint-action scenarios, humans and robots have to coordinate their movements to accomplish a given shared task. Lifting an object together, sawing a wood log, transferring objects from a point to another are all …

Deep LearningDeep Reinforcement LearningReinforcement Learning

Automatic Curriculum Design for Zero-Shot Human-AI Coordination

2025-03-10 · Won-Sang You, Tae-Gwan Ha, Seo-Young Lee, Kyung-Joong Kim

Zero-shot human-AI coordination is the training of an ego-agent to coordinate with humans without using human data. Most studies on zero-shot human-AI coordination have focused on enhancing the ego-agent's coordination a…

Tacit Coordination of Large Language Models

2026-01-28 · Ido Aharon, Emanuele La Malfa, Michael Wooldridge, Sarit Kraus arxiv

Large Language Models (LLMs) are increasingly deployed in multi-agent settings that require coordination without communication, from human-AI interaction to safety-critical scenarios. Humans often overcome the absence of…