paper-with-me

Papers

Modeling Others using Oneself in Multi-Agent Reinforcement Learning

2018-02-26 · ICML 2018 7 · Roberta Raileanu, Emily Denton, Arthur Szlam, Rob Fergus

We consider the multi-agent reinforcement learning setting with imperfect information in which each agent is trying to maximize its own utility. The reward function depends on the hidden state (or goal) of both agents, so the agents must infer the other players' hidden goals from their observed behavior in order to solve the tasks. We propose a new approach for learning in these domains: Self Other-Modeling (SOM), in which an agent uses its own policy to predict the other agent's actions and update its belief of their hidden state in an online manner. We evaluate this approach on three different tasks and show that the agents are able to learn better policies using their estimate of the other players' hidden states, in both cooperative and adversarial settings.

📄 PDF Abstract BibTeX arXiv:1802.09640

Code (1)

cts198859/deeprl_dist tf

Tasks

Multi-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

The Conditions of Physical Embodiment Enable Generalization and Care

2025-10-08 · Leonardo Christov-Moore, Arthur Juliani, Alex Kiefer, Joel Lehman 외 arxiv

As artificial agents enter open-ended physical environments -- eldercare, disaster response, and space missions -- they must persist under uncertainty while providing reliable care. Yet current systems struggle to genera…

Modeling the Mistakes of Boundedly Rational Agents Within a Bayesian Theory of Mind

2021-06-24 · Arwa Alanqary, Gloria Z. Lin, Joie Le, Tan Zhi-Xuan 외

When inferring the goals that others are trying to achieve, people intuitively understand that others might make mistakes along the way. This is crucial for activities such as teaching, offering assistance, and deciding …

Game of Chess

A Formalization of Kant's Second Formulation of the Categorical Imperative

2018-01-09 · Felix Lindner, Martin Mose Bentzen

We present a formalization and computational implementation of the second formulation of Kant's categorical imperative. This ethical principle requires an agent to never treat someone merely as a means but always also as…

Selective Deficits in LLM Mental Self-Modeling in a Behavior-Based Test of Theory of Mind

2026-03-27 · Christopher Ackerman arxiv

The ability to represent oneself and others as agents with knowledge, intentions, and belief states that guide their behavior - Theory of Mind - is a human universal that enables us to navigate - and manipulate - the soc…

Neural Amortized Inference for Nested Multi-agent Reasoning

2023-08-21 · Kunal Jha, Tuan Anh Le, Chuanyang Jin, Yen-Ling Kuo 외

Multi-agent interactions, such as communication, teaching, and bluffing, often rely on higher-order social inference, i.e., understanding how others infer oneself. Such intricate reasoning can be effectively modeled thro…